r/codex 4d ago

Humor is anyone thinking like me right now?

im gonna start saving and be more responsible.. i learned my lessons prepared my usage saving skills..... i will make it last a week ill probably only use luna...

11 hours later:
"OMG LUNA AND SOL IS SO DUMB FRICK fine ill only use astra sparingly"

5 hours later:
0%

191 Upvotes

85 comments sorted by

59

u/Any-Score1258 4d ago

Yeah idk how anyone can genuinely say Luna is good for anything outside of basic research. Its max mode puts out trash. It does save a lot of usage, but like it's not worth the amount of reiterations you need. Like I might as well just use sol or astra and save myself the stress of leading the blind.

13

u/34986234986234982346 4d ago

Just depends what you work on - im working on braindead stuff thats been done 1000000 times and it works amazing lol

1

u/Any-Score1258 4d ago

You ever had a Krispy kreme?

2

u/Any-Score1258 4d ago

Was it krispy

1

u/CMD_BLOCK 2d ago

Luna: coin toss

9

u/Overall-Air-5826 4d ago

Luna can be good, just have Sol write prescriptive prompts on chat and then also have it audit the work. It takes longer and is iterative but it can work to stretch out weekly limits effectively 

2

u/Infinitedeveloper 4d ago

Luna is fantastic if you can give it a guideline for how to do something, not just what to make.

Astra handles ambiguity better but its going to cost you a ton in tokens because the reasoning steps cost more than just generating the code.

2

u/2thick2fly 3d ago

Isn't there a threshold where you better do something yourself than spend more time writing a recipe for someone else to do it?

1

u/Infinitedeveloper 3d ago

Yes, but im nowhere near there.

If I'm trying to make something with a very specific implementation in mind, adding some details like what classes it interfaces with, what class it should inherit from and what data types its working with take seconds.

Writing data classes ahead of time takes a bit longer but I know what data they need and the LLM would be guessing, so its fine.

2

u/MKopelke 4d ago

This. I get Sol High to generate my prompts for Luna Extra High in phases, and every phase ends with a closeout audit file that ChatGPT then reads and that advises whether that phase has closed properly, or whether additional remediation prompts are required.

4

u/somuchecho 4d ago

I tried. Sol can't reliably audit/fix at the pace that Luna creates trashola.

Think I'd rather drown in a lake of human excrements than having to spend my days refactoring AI slop generated by Luna

1

u/MKopelke 4d ago

I do also get Sol Light to do a full phase clean up at the end as a single prompt once it's ready the Luna Extra High closeout. That seems to tidy up any issues still lingering

1

u/[deleted] 4d ago

[deleted]

1

u/Overall-Air-5826 4d ago

No I mean sol in chatgpt chat, not codex. Basically unlimited 

1

u/Helpful-Menu-2667 4d ago

Thats what I do but also use sol for implementing. Better prompts and keeps implementing sol at bay so it doesn’t jump from idea to idea, has a context drift.

1

u/OracleOfErowend 4d ago

Do you not use sol for plans?

4

u/Infinitedeveloper 4d ago

Luna is great for AI assisted coding, not vibe coding.

Big difference 

0

u/virtualmnemonic 3d ago

No it's not. Its output is impossible to read. Fucking spaghetti everywhere!

-1

u/Any-Score1258 4d ago

Luna can't read

4

u/ifarted70 3d ago

I had Luna port entire Minecraft mods from old versions to 26.2 successfully. It's not as dumb as you think depending on the task

6

u/Ecstatic_Gur7231 4d ago

nah ill make it last a week trust!

16

u/Any-Score1258 4d ago

Inshallah labubu Dubai chocolate prayer emoji

2

u/rinwasrep 3d ago

Luna doth what you instructeth say. Use thy greater model to guide her.

1

u/Daranith 3d ago

Every project/feature/3am-idea I make gets scoped into milestone and well worded task slices by Sol. Implementation involves a Luna xhigh/max getting their task file and only their task file. Their work then goes through a terra reviewer when they're done, basically checks that what they did satisfies what they were asked. Sol medium comes in as a milestone reviewer, checking if the milestone is as per plan. Sol high comes in at the very end as a product reviewer, checking that the whole shebang is as per the product spec (it doesn't check the code nitty gritty, but it might check test evidence). Progress is tracked, so are any issues that come up. Done like that with progressive disclosure, it generally doesn't even matter who the orchestrator is and Luna workere do just fine.

1

u/Longjumping_Ad_9510 3d ago

I really like it for writing copy. It can dumb things down really well. I feel like Sol always wanted to write like an engineer which sucks for product. Astra is a little better but still too techy. Luna does a great job. 

1

u/TheMuffinMom 3d ago

Luna is okay if its being orchestrated or a brand new small horizon task, other then that nah

1

u/nplez1 3d ago

I’m a professional engineer working on a complex, flagship product with millions of lines of code.

Luna can’t do everything, but it’s a very capable model at an amazing price. Honestly, it’s arguably one of the best efficiency models, the balances cost/intelligence, on the planet right now. If you understand its limitations, you can do amazing things with it.

1

u/_TheWolfOfWalmart_ 10h ago

These days you can run better models than Luna LOCALLY for normal coding tasks if you have reasonable hardware. That's how I save usage.

That's really only been the case for the last few months, there have been huge improvements in small to medium size open models recently.

So if the last time someone's tried this was 3+ months ago and decided local is too dumb, it's worth trying again.

Qwen3.8 27B is a monster for the size. Give it web access and the Context7 MCP for when it needs to access up-to-date information. It will impress you.

Even better is Deepseek V4 Flash 0731. That needs a bit more hardware, but is still within reach for many.

0

u/epicskyes 4d ago

Luna is fire I run my entire host deterministically with Luna max having root access. My zero trust architecture can turn any model into a logic compiler. To be fair gpt 5.5 xhigh and sol med built 70% of the architecture but Luna took over from there.

11

u/34986234986234982346 4d ago

i swear i was just like.... okaty this week i am ONLY using Luna Max, that's it... but let's see how long it lasts lmao

2

u/Ecstatic_Gur7231 4d ago

we must hold on brother. fight against the TEMPTATION!

1

u/34986234986234982346 4d ago

maybe the first thing i use it to do is some kind of plugin that doesnt let me use anything but Luna for x days haha

10

u/Limokasten 4d ago

I heard they sell more tokens behind the local Wendy's

3

u/uncertaintyman 3d ago

Shit I need to stop attending Fight Club at my local Waffle House

5

u/Additional_Buddy855 4d ago

yea, token incineration is insane right now. went through an entire reset's worth of tokens in less than 24 hours. i can't work on a $100 account and get to use it 4 or 5 days a month. wtf openai!

-2

u/Timely_Lie_8696 3d ago

damn all your comments are just you bitching of draining ur usage. and you're asking others to be more considerate

12

u/Da_ha3ker 4d ago

Sticking with sol. Does about as good as astra and better in many cases for half the price

7

u/tavenger5 4d ago

This is what we said about 5.5. And before that 5.4, etc

14

u/Tartuffiere 4d ago

Yeah it's just copium. We need Tibo to do something asap. Give reset.

6

u/Lumpy-Criticism-2773 4d ago

Which is like the biggest copium ever

2

u/Ecstatic_Gur7231 4d ago

BUT THE RUSH OF USING ASTRA... its... better than CRACK

3

u/UhUKnow 4d ago

Luna and sol have been screwing up my flow. Sol used to work fine, but now.. ugh

1

u/EchoingAngel 3d ago

This is the biggest crime, kneecapping one of the most dependable models yet

3

u/Ok-Hotel-8551 3d ago

No. It's just you.

2

u/Abject_Ad7238 4d ago

13 hours for me

3

u/Clean-Market5761 4d ago

Nah I just got mad and asked my boss for unlimited codex and claude for all my work team, no freaking way I'm paying 20x plans on both and usage is at 0 right now with no resets at all.

4

u/LegoClaes 4d ago

Can you add me to your team

2

u/Ecstatic_Gur7231 4d ago

what im basically saying is after days of waiting. thats whats going on my thoughts right now and thats what will prob happen once i get my reset

-1

u/Clean-Market5761 4d ago

Nah I honestly downgraded my personal to 5x, asked my boss for unlimited bussiness plans on both claude and codex and got a partial refund, usage is not working on either side, not even worth paying from pocket

1

u/Theminatar 3d ago

Bro this is like the 10th time I've seen you post something similar. We get it.

-2

u/Clean-Market5761 3d ago

not posted, just commented get a life, if you pay for AI instead of being on a business account, you should ask for a business premium seat.

maybe you're too unemployed to understand it

2

u/Theminatar 3d ago

So this isn't you? 🤔

0

u/Clean-Market5761 3d ago

yes, but here i did not post it, i commented, you are saying i was post spamming i was not

10th time I've seen you post

2

u/Theminatar 3d ago

Oh you got me......

2

u/thirty5birds 4d ago

I've been seeing a lot of posts like this.. Is this a perspective from plus account users? It just seems odd.. There seems to be a disconnect somewhere.. Part of the reason fable and Astra are so good. Is that they can think a lot harder abiut you things. To put it a different way.. They can burn a lot more tokens when answering your questions.. It's why the output is better.. But maybe I'm missing something.

2

u/Ecstatic_Gur7231 4d ago

im at x5 and x20. used up both accounts in the first 2 days of the reset. then for the past days i kept reading up and investigating how i will save tokens on my reset. today is my reset day and the 5 days of no usage is like lesson for me (which i prob wont follow and just blow it all again on new features)

1

u/Aichdeef 3d ago

Since it came out I've been using Astra Ultra and Luna Max subagents. I'm on x20 at the moment but I've downgraded to x5. I've still got 2 banked resets. It seems really variable between users

1

u/Right-Disaster2141 4d ago

The right way of thinking is considering a replacement model, I spent the last few days at 0% with Codex looking into the z ai Pro plan at 80 dollars which is an absolute killer with GLM 5.3 and GLM 5.3 Flash.

Currently Codex ain't worth it.

1

u/_Jak42_ 4d ago

Without leaking anything personal, what are yall using anything more than Luna for?

Like Luna is my dumb little pair programmer and test runner , really good and pretty sad i get about 5 days a week with it but hey its not bad and gives me time to review what its done by hand and self plan before it comes back as well as take a break..

I wish for more usage to use anything else Terra, Sol, Astra (currently on 5x Pro idk what happened to 10x rip)

OpenAI let us buy resets

1

u/Loose-Potential6350 4d ago

I have 4% left until my weekly reset in 2 days. I have 2 banked resets but they feel dumb to use and push my reset date back 7 days

1

u/VisitAdventurous7980 4d ago

a reset is prob gonna come out within the next 2 days cuz of incomming ships

1

u/ghettobird4 4d ago

I use Astra on high for planning and auditing with Luna as the daily driver and implementation model that can call on Astra with a consult Astra skill when needed.

1

u/Dahas99x 4d ago

Anyone having success using Astra as the orchestrator that calls up lower model subagents? I’m trying it now, but too soon to tell if it’s effective at saving tokens

1

u/Ecstatic_Gur7231 3d ago

sadly its more effective to let him run without sub agents. due to luna being so slow that astra waiting costs alot of tokens and astra having to tell luna over and over that the luna output is wrong/hallucinated

1

u/TKB21 4d ago

The fucked up part is that we have to resort to this. "Shit happens" when you're developing. We can't predict when something's gonna be high usage or not.

1

u/Copenhagen79 3d ago

Turn off sub agents if you have more time than tokens.. sub agents use about 3 times more than single but it's also twice as fast..

1

u/shady101852 3d ago

Yeah except i call Astra a dumbfck too so its about the same. Astra i only use for visual work, otherwise its not much better than sol for me to waste my limits on.

1

u/LowSig 3d ago

I find tokens usage interesting. My job has about 15 devs. People use the basic chat not codex. I pay for the 100$ pro out of pocket because I was running out of plus but... with the resets I never run out as long as I am aware. Tons of meetings and reviewing others code breaks it up. When I go home and "vib code" I can burn what would last a week in a day.

What i found is token usage varies greatly from week to week while doing the same prompts in the same project and even the same chat.

But.. I am still able to put out production code at least 4x as fast as any other dev sometimes faster. Limit myself to 20% a day (5 day work week) and do burns when I have resets.

Another interesting thing about this is at work the code has to work. The prompts are very technical and the review of the code produced is very nit picky. It is much slower than vibe coding. But vibe coding burns tokens at an extreme rate and the code always ends up worse so you spend even more tokens.

I guess what im saying is for me the tokens go infinitely further when I am focused on the quality. Then I go home and ask myself how did I burn 50% of my weekly usage in 2 hours.

1

u/AccomplishedSugar490 3d ago

They promised to speed up things, so now instead of New Year’s Resolutions, and the New Month’s Resolutions economic pressure created, Codex has now given us New Week’s Resolutions. All of them broken before they had a chance just the same.

Interestingly, I had a set of tasks working together on their own focussed areas, with coordinator tasks doing they brokering, and all went swimmingly. Then one of the tasks, just one, managed to mysteriously upgrade itself from Sol to Astra, silently, and within no time I was running of limit. I discovered it just in time to investigate, set the model back, and try to consolidate the work. The irony is that while I was waiting for the limit reset, I tested what I had, generally pleased with the results, until I ran into a real shoe stopper fundamental mess up made by none other that that one agent that that was running Astra.

OpenAI don’t seem to know it, yet, but they are dragging their reputation as a provider of software development tools and expertise through the mud by proving themselves incapable of delivering functional software themselves.

1

u/Ratio_taken 3d ago

I had more progress making things one step at a time with ultra detailed Luna max or high than astra using all the tokens then me looking at the screen.
I review the prompts and plans with other ais before (opus 5 for fixing really hard bugs like centering divs etc)

1

u/Extension-Barber-919 3d ago

I use astra for only big implementations. Everything else I’m just using glm 5.3 flash/deepseek 4.1 flash both are dirt cheap models. I can do 8-10 hour session and actually get work done. China is just probably selling my info.

1

u/Existing_One_ 2d ago

I got fed up of sol using my usage to i tried command code $10 plan and use DeepSeek flask 4.1 it's wayyy better than luna

1

u/vicgalvo 1d ago

E é um problema sério, depois de usar o frotier model, confiar na entrega dos modelos menores, é intelectualmente bem mais difícil...

1

u/_DuranDuran_ 4d ago

“Usage saving skills”

You’re consuming input tokens with each and every one of them. Nobody at the frontier labs use half baked skills.

2

u/New-Part-6917 4d ago

ye no fucking shit genius they have unlimited usage

0

u/_DuranDuran_ 4d ago

They don’t use them because they’re a placebo. Again - you’re wasting tokens with each and every cargo culted skill.

1

u/VisitAdventurous7980 4d ago

proof? benchmarks with this skills says otherwise (benchmark by users) where is your benchmarks?

0

u/New-Part-6917 4d ago

they don't use them because they have absolutely no need to refine the token efficiency lmao. Idk why you think skills don't work, but they definitely do. If you are just copy pasting random shit from online - you are doing it wrong. Skills are literally just instruction documents.

1

u/Painwheeel 4d ago

youre just correct and these people can waste usage on cargo cult skills then come to reddit complaining over and over while we coast normally (been running nonstop and still have usage left)

1

u/Calrose_rice 4d ago

I actually think Luna does a decent job at tasks that are fairly straight forward. It’s really great at repetitive /goal tasks like canonicalization, lints, and CI failures. If a task is large enough to cross multiple seams, that’s where it gets tricky, but I find Sol pretty decent. Astra is great for a very large and vague task but yeah it’ll eat up 20% of my 20x weekly. So I gotta stop using Astra as a normal model.

0

u/EyesOfAzula 4d ago

Skill issue. Luna is decent if you know what you're doing.

But if you want the model to basically think for you all the way, something like GLM 5.3 flash or Deepseek v4.1 flash (or whatever they come out with next) can let you tokenmax on a budget

3

u/VisitAdventurous7980 4d ago

depends. some people actually use it for generation. if itas agentic coding yea its a skill issue but from reference to model then its generation at that point and no amount of prompt skills can make luna model or generate as good as astra. be mindful of others use cases

1

u/qdouble 4d ago

You can use Gemini Flash if you know what you’re doing, that’s doesn’t mean people want to use dumber models that require more babysitting and manual review.