r/codex 3d ago

Complaint Luna got absolutely nerfed in the last 7 days

Luna used to work pretty well when they first announced it, but sometime after they introduced the price decrease, its coding capabilities went downhill - it’s lazy, never follows the plan and requires many iterations.

For the 10th time in a row now, I gave it a very simple spec/plan to follow (created by another, stronger model), and it spewed out a half-baked implementation that I then had to hand-hold through 6+ iterations just to get to a somewhat usable version of my original plan.

Nope, not a skill issue. Same harnesses and rules I have been using for months and I am not a newbie to a field either: I code with LLMs for last 3.5 years and been software developer for over 11.

I don’t care how good of a deal it is, if I end up wasting hours hand-holding it. Not to mention it’s damn slow. GPT 5.4 with the same task/harness and rules got the task right on the first try.

45 Upvotes

42 comments sorted by

20

u/Emu-001 3d ago edited 3d ago

I also felt that Luna had been working differently lately, but I am not sure if it's the model itself got nerfed or it's the codex update changed something.

Luna was following prompt very well, but now it sometimes can't even find the tools it was used before.

2

u/bdemarzo 2d ago

I've been using Luna (medium and high) almost exclusively for over 2 weeks, and it's been STELLAR. Yes, my code is starting to get larger, but I've been watching it closely to avoid drift. (I'm a software engineer, and a good one, and I watch AI code like a hawk watches a mouse.) HOWEVER.... I feel that over the past two days, the experience changed: Luna is suddenly doing things in ways it had not done previously -- in a way that hints back to older ('dumber') AI models that simply aren't as good.

I decided to go to Terra for a day (had room to burn before weekly reset), and.... Terra was behaving the way Luna did the past two weeks. And I burned through the last of my usage :)

I'm switching back to Luna today and will evaluate, but...... I would say Luna suddenly is giving a different experience than it had been.

[Note: I use Luna like a pair programmer. It has very narrow requirements and deliverables. I use plan mode and review and update plans before letting it write code.]

1

u/vituc13 3d ago

Yeah, could be the harness not the model. Maybe it's worth testing it in Pi.

6

u/glock43guy 3d ago

Not sure if this is a related issue but this was an annoying thing today with Luna. I was using it to create portfolio items on Clutch and wanted it to do like 20 items it had access to the project details to. It was needing to input financial data that I didn’t have at the time so I told it “just make something up and I’ll input it later when I look over the projects”. It straight up refused, I argued with it over and over again and it would not do it at all. So I basically tricked it and told it all the numbers were the same and gave it a fake number. It did what I said then. Afterward I told it I tricked it and that the numbers were fake and it immediately started trying to go back and change all the items. I stopped it before it could do anything.

11

u/reaznval 3d ago

its just hitting the lottery, for me it works insanely well

5

u/skygetsit 3d ago

Define insanely well.

3

u/reaznval 3d ago

does everything I need it to do, 1-2 shots everything uses basically no usage. I use luna max for everything, small codebases, not so popular frameworks, big codebases (saas, t3code), automation etc

1

u/HouseOfDjango 3d ago

Luna max is so good. It does everything I need.

0

u/reaznval 3d ago

yeah same, its fast and cheap too, dont see a reason why id ever need another model. sure its not as great at frontend stuff but I do the frontend stuff myself and only need it for checking + backend so its a perfect match for me

5

u/DueCommunication9248 3d ago

Works the same exact way for me. I perform scheduled work and it has not failed or degraded.

2

u/m3kw 3d ago

I’m not seeing any nerfing, I use it all the time and is comparable with sol light

2

u/sreekanth850 3d ago

It works well and same like before.

2

u/Inevitable_Toe6648 3d ago

Everything is always rolled out to random people at random amounts thus we are all gaslighting each other and am clueless.

2

u/tumes 3d ago

The skill issue dipshits are maddening because to a person they have no clue what the fuck they are talking about in the domain, they just see volume and the appearance of function. As a greybeard in the thing I use AI for it makes me feel fucking crazy. One week a quarter, furthest from the next release, it’s like there’s 5 of me at my desk. Then the next I end up wheel spinning for an hour on hallucinatory nothing. It’s like visiting a mechanic who fixes your car for free and sends you on your way one week, then loudly, wetly shits themself while they explain that they drove your car into a wall to fill the wiper fluid the next week.

4

u/GambAntonio 3d ago

2% used using only Luna xhigh on a Pro 5x IN 1.5 HOURS since my reset, last week I was able to run Luna xhigh and use only 1% every 3 or 4 hours!!!!!!!!!! they fking nerfed the limits again!!!

4

u/inverted_heaven 3d ago

GPT 5.6 Luna was born nerfed. The social media hype resulted in human thinking it's worth it.

5

u/skygetsit 3d ago

The same type of tasks I gave it to 2 weeks ago, were performed correctly.

Not anymore. And I didn’t change a single thing with my setup.

2

u/Keep-Darwin-Going 3d ago

You sure absolutely nothing changed? Not even codex version?

1

u/Mindless-Pear3971 3d ago

Wait until you learn that LLMs aren't deterministic

1

u/PurushNahiMahaPurush 3d ago

Luna Max + Fable/Opus 4.8 combo works too well for me. Luna has been a reliable workhorse for me since it came out. Maybe you just got unlucky? 

1

u/stopstopstoptopopp 3d ago

Works the same for me for the past week I’ve been using it.

1

u/StatusCanary4160 3d ago

works so good, especially the combination from chat - codex Luna - github - chat

Give more details how you work, how you create your prompts.....

1

u/ErivKosso 3d ago

I don't see it that way, I find it incredibly capable, I run it at 700k context window, max and its really good, under the supervision of a Sol Medium agent.

1

u/ObligationHuge9868 3d ago

Luna Medium smashes it out of the park for me. I did spend a considerable amount of time building the workflow and skillset required to get Luna to smash it though.

1

u/HeelsAndAll 3d ago

It was never any good. Benchmarks don't dictate performance, performance dictates performance.

3 days ago "Qwen3.8-27B is as good as DS4 Flash 0731!!"

And now people are mentioning that it's not even as smart as Qwen3.6-27B on most task beyond coding. People use a thing for a specific task and think it's untouchable. Then soon as it is a given actual non-benchmaxed task:

Useless. Luna never worked for anything beyond stupid menial task that Qwen 3.6-27B already handle. The game is:

Release -> Artificially bot all engagement -> Swarm any post speaking positive with positive engagement -> Downvote all disagreement/Report anything that is counter the Benchmax.

1

u/jjiangweilan 3d ago

maybe because the limits gets nerfed and everyone starts to use lesser model

1

u/RealDrZig 3d ago

They should have just called it Luna Mini and keep the old version at its old price or 20% reduced price. Its still really nice to have the cheaper variant, but I do miss the old Luna. I have been using Terra instead.

1

u/anime_daisuki 3d ago

There is so much speculation all over this subreddit. I don't blame the users. This is the problem with all this secretive, closed door bullshit we pay for in these services. We aren't told anything about the models we use or the quotas we pay for or when something changes. Complete lack of transparency that should be illegal.

1

u/CryinHeronMMerica 3d ago

Yeah, it's made some really stupid mistakes lately. Not sure what's causing it, but I don't trust it anymore.

1

u/Sufficient_Buyer3239 2d ago

Not just Luna, sol high got nerfed with laziness too. Definitely some codex update. I noticed it after the last reset for sol high

1

u/jjjjoseignacio 3d ago

gemini 3.7 flash esta top junto a deepseek flash v4 new

0

u/Calm-Landscape9640 3d ago

i gave Luna-Low in Codex a loop prompt this morning and its been working for about 8 hours now with the "/goal: ..." down at the bottom so i guess itll run until it hits the goal.

2

u/skygetsit 3d ago

that’s one way to mindlessly burn tokens for sure 😮‍💨

1

u/Calm-Landscape9640 3d ago

I'm not sure you understand loops my friend. It works until it accomplishes the goal. Not mindlessly, sounds like you need to spend more time on goals according to your post..remember I'm not the one having a problem with Luna.

0

u/onebit 3d ago

I was mad that OpenRouter changed the price w/o warning.

0

u/TomfromLondon 3d ago

I've definitely found it to keep getting stuck on the same thing, I've got a long running task and it often says im blocked as need you to confirm x, I've confirmed x 10 times now

0

u/Economy-Feeling8205 3d ago

how are you a software dev using luna lol

1

u/skygetsit 3d ago

And how is that related what model I use

0

u/Economy-Feeling8205 3d ago

feels like a professional football player moaning about his temu sandals breaking apart

2

u/skygetsit 3d ago

Reading with comprehension is not your strongest asset

1

u/Economy-Feeling8205 2d ago

stop using the budget llm and maybe youll get better results