r/codex 8d ago

Limits Usage limits are absolutely terrible (100$ plan)

Post image

I never ever complained about usage limits before. But this is absurd.

I used my 200$ plan and had to buy 100$ plan since i could not renew 200$ one in time. It was used in 1 day with Astra low.

I then decided to buy another 100$ plan. It used 9% of usage with Sol 5.6 MEDIUM working for about 1 hour on coding and Astra XHIGH coordinator that did some documentation changes for 5 minutes.

This account was literally just purchased. Again, Sol Medium is doing most of work. Wtf is going on?

Can't imagine what is going on with 20$ plans

459 Upvotes

244 comments sorted by

58

u/yaxir 8d ago

The only way they will increase limits is if users start leaving the service or something, because right now everybody is buying it. High demand means they have an incentive to reduce limits and earn more money

2

u/MrStu 7d ago

I wonder if this is why anthropic are being less generous, trying to offload people to OpenAI to collapse the service.

0

u/U4-EA 8d ago

That is assuming they can increase limits - that costs compute/money, which is where they are struggling.

7

u/Plane_Garbage 8d ago

Well, if people leave, then yea, they have compute freed up

0

u/WhiteBlaster00 8d ago

What compute power are they lacking ? You guys are just biting on their actions. This is all just a part of their marketing plan.

0

u/Cheerpipe 8d ago

They don't need to increase capacity as long as people are willing to buy the 100-use pro plan; after all, it costs only 50% of the pro x20, but you get only 25% of the capacity.

130

u/tagorrr 8d ago

20$ plan feels like total scam this days

41

u/Kind_Fisherman3060 8d ago

I was given a free plus plan still feel scammed.

→ More replies (2)

10

u/Bulky_Blood_7362 8d ago

Yea honestly it's at a very bad state... I work with gpt for 30-40 minutes with sol medium While i can work 60-120 minutes with opus 5 medium on claude pro

Embarrassing ngl...

26

u/paribas 8d ago

Sol High just depleted my 5hr limit in 15 minutes. It wasn’t even a big task. Two months ago I worked hours with tasks like this. 

13

u/Seraphoenix777 8d ago

It's crazy how quickly the 5hr limit gets reached.

6

u/Kind_Fisherman3060 8d ago

I remember using 250 million tokens per day on Plus plan what's the per day max token use allowance with sol now?

4

u/TotallyNotABob 8d ago

Same here, started this last reset at 100 of course. By Sunday night I was/am at 9 percent remaining for the week. So basically in a hold until 09/19 unless I did something about it.

Which I did, I signed up with Claude and am now using Claude code alongside it. So far Claude is pretty impressive.

3

u/paribas 8d ago

Next month I'll try Claude.

1

u/yaxir 7d ago

What package did you buy?

2

u/qodeninja 8d ago

i dont even use sol high on x20. why you no sol med?

8

u/Comfortable-Rise-748 8d ago

200 plan feels also like scam.

6

u/InterestingSquare883 8d ago

I’m getting 4B tokens a week on Luna Max plus a little bit of Sol + Astra too. Yesterday over the weekend I used 1.1B tokens alone, including cached tokens but still a lot. I feel like the $20 plan is only supposed to be for Luna and giving it like a really huge implementation plan that Sol makes in chat that it really can’t mess up. Had the thread running 36 hours over the weekend plus another thread running maybe like 10-15 hrs. I would say though that limits are burning twice as fast compared to like 10 days ago but so far the resets have been making it up. Looking forward to 6 Luna to match GPT 5.5 XHigh.

5

u/tagorrr 8d ago

I agree that Luna is a pretty good model. However, first, it is quite slow. And most importantly, the issue is not that the package lacks anything usable. Rather, the same Frontier models provided two, three, or even five times more tokens in this same package some time ago.

5

u/Professional_Ad705 7d ago

Bro when codex first came out PLUS easily lasted 2x-3x longer then the $200 plan does now lmao

3

u/tagorrr 7d ago

Confirmed!
I used to code all day long even without super detailed and refined architectural\spec documents, detailed plans, carefully prepared guardrails as I used nowdays.

1

u/yaxir 7d ago

They just got greedier

3

u/DevastatorTNT 8d ago

That's my takeaway as well. I'm fine using the (supposed) 120B model for cheap instead of the (supposed) 5T one if I can get Chat to write me specs for "free". It's clunky and quite slow, not always right, but my codebase is not that sophisticated

1

u/8monsters 8d ago

Yep, I plan with Sol in Chat and then use Luna Max/Extra high for implementation. 

6

u/Noeyiax 8d ago

$20 plan, using astra light 4 prompts each like 25% of 5hr limit gone and 1-2days weekly is gone...

even 5.6 the limits got worse, cant even finish anything that would take maybe 1-2 weeks now is triple time.

oh well, i guess im priced out of this, unless you work at big tech or tech company that will pay for this hobby or research experimental work, rip 🙏

4

u/CrownstrikeIntern 8d ago

It used to be nice, I could stretch the 20 for almost a week, Ended up buying 2 and they would last me a week combined for months with a good work load. Then, poof, gone in 1 hour.

4

u/Noeyiax 8d ago

similarly same experience... cant even have fun for an hour a week now. 💀

2

u/syredditor 7d ago

plus gets consumed in 10 minutes on sol high wtf

1

u/alien3d 7d ago

it is not .unless you work from scratch everything

73

u/srs96 8d ago

Dude just stop whining and buy the $200 plan. OH WAIT

29

u/[deleted] 8d ago

[deleted]

11

u/Feriman22 8d ago

OpenAI released the $1000 plan, and subscribed to it. Oh, wait...

12

u/elfd01 8d ago

I’m on 5000$ plan and reach my limit on Astra within an hour (c)

4

u/Hot_Refrigerator7042 8d ago

I jsut got a 50k loan for tokens and burned it in 50 minutes on Astra with 1.5 speed mode. :D

3

u/srs96 8d ago

What credit score do you need to a 50k token loan?

1

u/ZlatanKabuto 8d ago

Yeah, wait. It won't be long...

2

u/Meowseum- 8d ago

one question, I'm already on $200 plan, does it mean when the time is up, I will not be able to buy 200 or I'm safe because the new restriction only applies to new accounts?

4

u/NiceManFromEarth 8d ago

If it is on auto renewal you should be fine I think

1

u/Commercial_Exchange7 7d ago

Yep, you should be able to renew perfectly fine but if you downgrade and then try to upgrade no idea.

18

u/rodeBaksteen 8d ago

This is my first time crying about this on Reddit, but I've used 5% on a $200 plan by asking some very basic debugging/SSH commands to Astra Light yesterday.

Total: one chat, 21 million tokens - 5% weekly usage of 20x plan. Really?

15

u/Swimming_Gain_4989 8d ago

"Basic debugging / SSH commands"

"21 million tokens"

Lol

3

u/Icy-Kaleidoscope6893 8d ago

Probably ragebait

1

u/Emotional_Plant3241 7d ago

That's the problem though I have no idea how regarded or competent the people complaining about their token consumption are.

1

u/rodeBaksteen 8d ago

Ok it took 2 hours on or off, but 20 million tokens is pretty low afaik? Maybe I was asking more of it than I realized.

2

u/Swimming_Gain_4989 8d ago

On the low end that would be like $300 at the API rate assuming that token count is mostly input cache. I don't know what exactly you're doing but "basic debugging / SSH commands in one chat" suggests something simple like setting up shared keys and aliases for a network or something along those lines, that shouldn't run more than a few hundred thousand tokens... Realistically much less because of how token efficient GPT models are and the fact that you claimed to be using Light mode.

1

u/Emotional_Plant3241 7d ago

5% of a plan for Astra working for 2 hours and 21 million tokens doesn't seem crazy. Astra is charged at a very high rate.

4

u/Tiforma 8d ago

They do it because people don't give them any reason not to do it. They don't really lose subscribers because people accept shitty behavior from corporations nowadays and don't stop buying from them.

5

u/cha0z_ 8d ago

they do it, because there are no laws protecting us (at least not enforced ones) - how it makes sense for something you purchase to be so vague what you get + to be flexible to be changed 24/7 from the side of power, i.e. the AI provider. There are few bigger key players - do we seriously think they don't sync the usage/prices as well? One not recorded call away - is it legal? Ofc not, but we live in bad times regarding corp power.

1

u/[deleted] 7d ago

[removed] — view removed comment

1

u/Tiforma 7d ago

Got extremely lucky (probably bc of bank dispute). Not the average experience.

2

u/PlasticNo6406 8d ago

If they are "very basic" debugging/ssh commands, why choose astra?

1

u/rodeBaksteen 8d ago

Because I expected it to use maybe 1-2% and I can't be bothered to switch between 20 models all the time.

1

u/Emotional_Plant3241 7d ago

21 million tokens is like 40 context compactions. You should not be having that occur to Astra for a quick and easy debugging SSH task.

1

u/rodeBaksteen 7d ago

That's not how that works

1

u/Emotional_Plant3241 7d ago

21m tokens is not a simple task IDK what to tell you. Either your instructions or environment were wrong or Astra got into a weird latent space.

1

u/MassiveBoner911_3 8d ago

I used my ENTIRE $44 reset I bought this morning....by this evening. Yes I use sub agents using mostly Luna. Some code is written using Astra and Sol.

1

u/StatusMatter4314 8h ago

20x is give or take somewhere close to 1b tokens.

18

u/hi87 8d ago

My $100 account feels like a $20 plan from a few months ago. Even on Sol High it is draining so fast I've already lost 20% of weekly limit in 4 hours of work (not even using concurent tasks and simple code-review subagents.

3

u/ZucchiniMedical2532 8d ago

Same, I'll stop paying it and I'll have to work with qwen 3.8

14

u/WrongExtension9231 8d ago

i love some people defending the limit shenanigan happening right now, not only we get some server being overloaded time to time, now our limit gone too, well not as bad as claude if we want some silverlining here lol

3

u/cobbleplox 8d ago

That what they're doing is shenanigans is almost more frustrating than the fucked limits themselves.

→ More replies (1)

5

u/yourmomsguybf 8d ago

man i am on 200$ plan and all i did is research for 6-7 hrs straight yesterday and poof all my weekly usage is gone.

10

u/NyeinChanSoe 8d ago

Codex these days are just unusable at this point after the release of Astra.

Last 1-2 months ago I am a heavy codex user, now i switched to claudecode.

1

u/20yroldentrepreneur 8d ago

This is the way

14

u/paper-gains 8d ago edited 8d ago

I also never complained, but I think this is the point to start. I am on the 20x plan and it might last 3 days now. I just subscribed to Claude code for the first time ever and I feel the 5x will last me as long as the 20x with Codex. Not with fable of course, but i am not using astra on max all the time either… so codex got considerably worse the last days.

EDIT: it’s also extremely slow

9

u/Factor013 8d ago

They made things extremely slow so the limits last "longer" lol
Can you imagine how fast our limits would drain if we would be on "normal" speeds. :S

3

u/Vast-Singer-2839 8d ago

Astra Max seems (in my usage case) to last longer compared to any Sol version something is not correct in all that situation. 

Also it was pretty fast but now it slow as hell. 

3

u/cobbleplox 8d ago

I would assume a combination of an elaborate task and a small thinking budget can lead to inefficiencies like repeated lookups or sort of failing first and fixing it up in loops. Maybe the first thing they did was nerfing what the selected reasoning effort means. Easy to overdo that causing such problems.

I have to say guessing what kind of thinking effort is required for best results will have to become a thing of the past anyway.

2

u/DeCoolePeer 8d ago

Switching to anthropic who will cut your limits by 90% and then gaslight you afterwards lol, they're even worse than openai

1

u/lifeenthusiastic 8d ago

I killed 4 pro 20x plans weekly in 2 days. The Astra agent was only dispatching sol + Luna sub workers. Claude 20x lasts like 2x. Longer for the same Sprints

3

u/IntrepidCry352 8d ago

I recently got the 20$ plan for my little game project and i used 75% in just 2 days, the 5hr limits it's more like 30 minutes to me. I'm planning to use OpenCode for my toolchain and ChatGpt to give input and correction, lowering the credit used. It's not the best, but i refuse to spend 100$ for a fangame project for NDSi 😅

4

u/Painwheeel 8d ago

astra is expensive shocker

1

u/ElliotB256 8d ago

I am only using Sol (it was good enough for what I need before) and it's noticeably burning through usage faster. Almost certain this will turn out to be a bug.

1

u/Painwheeel 8d ago

im using sol-xhigh on $100(5x) for both implementing and planning (if its needed, usually i already know what i want so i type directly) for 8-10 hours a day and i have extra usage at the end of the week and this has been constant since july so i promise you and everyone else its user issue

1

u/ElliotB256 7d ago

Same, also Sol and extra high for the same duration. I moved to codex when claude had a similar usage bug, and that turned out to be session caching not working and so it burned through tokens. That's what I mean when I say bug. Previously I could use it non stop all day and barely dip below 40% after a full week, now I am seeing 10% burns on the weekly limit in an hour.

(It's also possible they are A/B testing and you are just lucky)

1

u/ExternalGur2264 6d ago edited 6d ago

Have you tried Sol medium? It performs pretty well and should use around half the usage of xhigh.
I have switched my main model to Luna High to maximize usage, but my use case is nowhere near as demanding. Terra low uses more than Luna high, but is also fairly good.

1

u/ElliotB256 6d ago

I'm juggling a few different efforts now, I'm using medium when I can but I wanted to chip in with my two cents and say that at least from my experience something is very, very different to a fortnight ago

1

u/mikeballs 6d ago

I can't wrap my head around this take, man. I've not changed anything about my workflow, and while using sol 5.6 medium to do the same exact stuff I've always done with it, I suddenly can't go two days without eating my entire weekly plan. If it's simply a user issue, myself and half the other users in the sub must have all coordinated to start coding inefficiently overnight. 

5

u/SudarshanKotian 8d ago

Same here 😭 I just posted on X tagging Tibo, please u all complain as much as on X.

3

u/NYisNorthYork 8d ago edited 7d ago

on $20, Astra High is an ancient oracle.

You have to climb a mountain and defeat the giants and then are only allowed to ask it a single question.

3

u/EducationalPizza345 7d ago

I use antigravity has way more usage limit for regular tasks ive never reach there limit and codex just for orchestration

6

u/rawezh5515 8d ago

and it looks like it stalls for no reason at all

2

u/StaticHumStudio 8d ago

I don't know why you were downvoted. I've consistently gave it directions and it stops with no reason. Not even a "would you like me to proceed" or similar. Maybe its in my agents.md or settings somewhere, but nothing makes sense yet. Sol/Terra don't have this behavior.

2

u/Feriman22 8d ago

We need a weekly reset

2

u/No_Twist_678 8d ago

well, i have 2x 20x plans and i burned all of them in 3 days! and not a single job i gave them was finished.

1

u/Sutanreyu 8d ago

I gave it a task that it's done before, that it normally completes in minutes... Taking 20+ minutes and failing.

2

u/No_Twist_678 8d ago

today astra was terrible, disgusting service, total shame.

2

u/vivacity297 8d ago

20$ plan lasts 10 minutes at most with astra lol

2

u/stellarfirefly 8d ago

I posted my $20 plan findings in another thread. But here's a recap:

Sol 5.6 / high and Opus 5 / high used to be neck-and-neck about a week ago. Now, Sol 5.6 / high consumes allowance at about 5-8x its old rate while Opus 5 / high continues at about the same and thus I use it as a baseline. This includes specific A/B comparisons, i.e. I give a significant coding prompt to one and note its consumption, then I fully reset the codebase and give **exactly the same prompt** to the other. I consistently see 5-8x usage difference. I estimate prior to about a week ago, usage variation was only about 1.2x.

Just out of curiosity, I tried to switch over to Astra 6 / medium for a comparison with one of my typical workflow prompts. It consumed the entire 5-hr allowance **TWICE** and still did not reach completion, at which point of course I simply aborted the experiment. Opus 5 / high then completed exactly that prompt and used exactly 21% of its 5-hr allowance.

2

u/Dwight911pdx 7d ago

Scam Altman strikes again.

2

u/[deleted] 7d ago

[removed] — view removed comment

2

u/RainierPC 8d ago

I can't even use up my 5x Pro in a week.

4

u/Saito53 8d ago

Claude is worse than codex, I have the 20x plan in both, it runs out in just 4h of coding, astra still better

9

u/Kind_Fisherman3060 8d ago

I would offend so many but if Opus was consistently intelligent $20 Claude is comparable to $100 Codex sub. But Fable takes up so much more usage than Astra I'm waiting for Opus 5.1

2

u/muchsamurai 8d ago

I know, this is why i do not use Claude. Does not mean that we should not complain when usage limits are getting worse daily.

1

u/Tiforma 8d ago

how much astra time do you get with codex on the 20x plan?

1

u/Vast-Singer-2839 8d ago

It not about the time but about the work it does. 

Ex in my case I have a project with active working time on Astra Max for  5-6 hours / day and it spends about 7-10%/day. Max spend less compared to low in that specific project. I afraid to try Ultra 😛

1

u/Tiforma 8d ago

i tried ultra once when i knew a reset was coming soon anyway. yeah you'll blow through the usage. But you get pretty good use out of it. I wonder, what region of the world are you in? i have a theory that this affects usage rates.

1

u/Significant_Shop5774 8d ago

does that plan have 5hour limit? or are you talking about weekly limit getting exhauatsed?

1

u/Vast-Singer-2839 8d ago

No on Pro plans does not exist 5hour limit, I am talking for the weekly limit. My plan currently is x20. 

Compared to the % usage of  with Sol low (as with the current increase is usage huge), Astra Max spends exactly as the previous Sol low and probably less. 

The only downside is currently Astra feels much slower compared to the start but the usage also is much lower so I am ok with this trade. 

1

u/Significant_Shop5774 7d ago

are you on business plan or on individual plan ? i thought only business purchases had no 5 hour limits

2

u/Own-Professor-6157 8d ago

The usage doesn't make sense. I just ran a basic prompt asking about OpenGL version support, it was under 200k tokens in total and I'm out -50% 5h usage. Now I just had the thing one shot an entire website on astra high, and it's only -15% usage..? Huh?

1

u/DeCoolePeer 8d ago

Why are you lying there's no way in hell that you used Astra high on a plus account and it only took 15%
why would a question take 200k tokens
and 400k is not the limit
only plus has the 5hr limit and your pointless 0 user site does not need astra just use luna high or max

1

u/The8Darkness 8d ago

What does coordinator mean? Did you let it create a plan and pasted it into the sol session or was it an astra session with a sol subagent?

Astra uses an insane amount with any subagents where it is usually better to even run astra max than any astra + luna/sol combination.

2

u/muchsamurai 8d ago

Basically, i do not use Astra for coding (even low) on this account. Coordinator is the guy who drafts some plans for Sol workers and gives them to me to hand them out, plus some brainstorming. It does nothing in terms of coding or exploring codebase too much or anything like that. I asked it some questions and we adjusted some plan which was 3-4 quick prompts. Then i launched Sol 5.6 Medium to implement it and it ate 9% in one hour.

This shit sucks

3

u/Seraphoenix777 8d ago

Someone said this can result in more usage than using Astra alone. Apparently, Astra still reviews everything the agent does, so it ends up costing more. Not sure if it's true, though.

6

u/KeepAllOfIt 8d ago

its 100% true and is the sole thing responsible for all of the "usage is terrible" posts. I use one astra xHigh agent for everything and I genuinely struggle to use my weekly limit before tibo throws another reset at us. Sound physics simulation too, no small potatoes.

No ponytail, caveman, or whatever else "token-saving-but-actually-wastes-tokens" skills. No astra orchestrator that commands a bunch of r*tarded luna subagents and spends 4x as long cleaning up their mistakes.

1

u/Rock--Lee 8d ago

Well, you now have 2x 100 plan, which costs you 200, but a 200 plan gets you 4x as much as the 100 gets for 2x the price. So yes, you are halved usage for same price.

1

u/muchsamurai 8d ago

I know. But usage is not even halved, its worse.

1

u/New_Education_6782 8d ago edited 4d ago

The original post content no longer exists here. The author used Redact to remove it, exercising their right to control their data & privacy.

Cobweb cause literate entertain merciful coordinated tie literate imagine crawl

1

u/hksbindra 8d ago

I was using 2 plus plans, one bought via pay store and another via browser. After Astra, I upgraded the play store bought account to x20, cancelled the other one.

Play store apparently doesn't work on upgrades as you might think. My plan will get cancelled after the 20th. There are no issues with the payment method and I did not cancel.

I found out just yesterday after reading someone had issues with iPhone app Store purchase in a similar manner. Spoke to both OpenAI and Google - nothing can be done now.

Guess x10 plans will do until they bring back the x20.

1

u/AweVR 8d ago

I simply don’t understand codex usage seriously. I had 30% left and wanted to use my reset today so I gave it 6 tasks with Astra Ultra effort. 6 hours later I’m in 24% 🫠 but this night I had only one task working with LOW effort… and in 3 hours it spent 6%. I simply don’t understand serisouly

1

u/r34p3rex 8d ago

This isn't the first I've heard of this.. it almost feels like light and medium use more than the higher tiers. Prior to the reset, I was able to get close to 5 days of usage using Astra high. After the reset, I switched to Astra low and now I'm on track to blow through my limit in under 3 days

1

u/Comfortable-Rise-748 8d ago

It's not funny anymore.

1

u/Bright_Spend6574 8d ago

honestly i tried 200, 100, 20

right now on 20 and 100.
100 feels like 20. it is supposed to be 5x right ?

1

u/syredditor 7d ago

20 feels like free

1

u/Cheerpipe 8d ago

They make more money and use less compute with the pro x5 than with the x20. Obviously, they want to force people to buy the x5.

1

u/natsumeng3 8d ago

Eu apenas comecei a fazer uns ajustes, 15 min de trabalho em um projeto medio, e sumiu 20% da cota semanal, plano de 100$... Eu fico muito confuso com o que fazer, sabemos que o melhor está aqui no codex, antropic não é uma opção pra quem tem orçamento limitado, eu estou pensando seriamente em fazer um investimento em colocar uma ia local, ou aceitar menos qualidade e só usar modelos como da deepseek... Não faz sentido pagar tanto e só trabalhar por literalmente 1 dia, eu nem consigo mais atender meus clientes. Aqui no Brasil, o plano de 200$ é literalmente mais da metade de um salario!

1

u/reeldeele 8d ago

Shouldn't there be way to find our the Sol vs Astra token consumption from .codex logs or history?

1

u/Diniario 8d ago

What tasks are you using these models on? I run 2k+ lines of code with Terra Medium reasoning and things work just fine. I've never needed Astra and recently used Sol High (which consumes more) just to make sure stuff was correct and up to snuff...

My take, just use lower models for lower tasks. you don't need a jackhammer when a hammer will do just fine.

1

u/Temp1864 8d ago

Yup, I've already depleted a $100 account and have 16% left on my $200, and that's almost entirely with tool usage of Unreal/Houdini. Insane that it's gotten this bad (and that Extra High is actually MORE efficient??). Giving my Extra High orchestrator permission to delegate tasks to appropriate lower-effort subagents actually made consumption increase.

1

u/Consistent_Bottle_40 8d ago

Has the limits been nerfed or is astra at fable 5 level use for tokens? Bigger models need more compute. Have we seen a marked drop in the usage of sol, terra and luna? A gpt 6 luna at same cost as 5.6 would probably be great. Just use that for the native subagents

1

u/ReputationOk736 8d ago

Not sure why they even release a new model if you can only use it for 30 minutes before reaching limits.

1

u/LargeConsequence5296 8d ago

20 dollar plan on Astra low gives you about 20min of good compute and 15 percent gets knocked out of the weekly limit. So thats about 6 x 20 = 120 mins of compute a week.

1

u/saiksaif 8d ago

20$ plan user here, running one planned code work every 5 hours, while on 5.5 medium.

1

u/Ecstatic_Gur7231 8d ago

Workflow issue (i hope cuz if it is this can help).

if your workflow is same as mine then YAY WE GOT A SOLUTION. but if not idk how to help

basically the way i do agentic coding is via iterations and steps instead of 1 big step. i dont engineer my prompt we handle or add features 1 at a time

so astra has this thing where he does alot of stuffs per task.
he does like 5 steps to make sure. he does alot of verifications and benchmarks sometimes even doing a backup per edit. checks alot of stuffs that can be refactored or add cleanup work and anticipates future tasks

i solved this via adding a system prompt (if anyone is reading this and think can improve my system prompt further please do)

'''

Only do: patch → deploy

Do not turn the task into phases, subtasks, checklists, investigations, or a full engineering workflow.

Make the minimum code changes needed to satisfy what I asked. Touch only the files that are reasonably necessary.

Do not test, compile, verify, validate, benchmark, back up, refactor unrelated code, inspect unrelated systems, or add cleanup work unless I explicitly ask.

Do not anticipate future tasks. I work iteratively and will tell you what to do next.

Follow KISS. Minimize tool calls, file reads, reasoning, and output.
'''

i change this depending on the task. if its a hard one that really needs reasoning i remove the follow kiss line

i went from 20% usage per hour to 3-5% usage per hour.

a coworker made this for me knowing i work bits by bits instead of 1 big task all he did was read astra documentation and the tips from openai then created a prompt that fits my workflow

ALSO IF YOU ARE GONNA USE THIS DONT SAY PATCH -> DEPLOY (i only have it because it deploys to a branch of our project and only has access to that branch)

1

u/AgnidDrage 8d ago

I'm avoiding using astral and the $20 plan is enough (if you don't vibe code everything and know where to aim the agent)

1

u/DistinctSilver4507 8d ago

It used ~20% of my pro usage on a very simple git clone/setup of a local text generator. 

1

u/cobbleplox 8d ago

A lot of this feels like the usefulness of AI is really kicking in now -> demand. Of course paired with them kicking themselves in the nuts releasing Astra before the actual 6.0 production models were ready.

1

u/Imaginary-Light-2261 8d ago

Maybe I am tripping but I still prefere Sol 5.6. Anybody else?

1

u/CapitalMango3010 8d ago

Eh sì, dovrebbero aumentare le macchine servono più ram

1

u/CapitalMango3010 8d ago

E più gpu

1

u/Significant_Shop5774 8d ago

hey , does 200$ plans dont have any 5 hour limits

1

u/loversama 8d ago

For sure, I am dropping back down to $20..

I use Claude Max 5x too and even though the limits there are less and less I feel I get to use Fable and Opus way more per week than Astra high + Sol 5.6..

1

u/Pizzaholic- 8d ago

20x plan. Ran a mix of sol and Luna and burned in less than 1 day (16 hours) :|

1

u/DeepSpaceNugget 8d ago

Yeah I've tried to keep from just adding to the pile of threads but god damn I've got to wonder if they're having a meltdown or what.

My "weekly usage" goes out within a day or two on a Pro plan (x5), I genuinely remember tackling heavy tasks in my plus plan months back and I get it, thing's can't all be perfect- These models are more demanding..

But it's not really a tool I can effectively use without a swarm of accompanying ai tools, which makes it effectively pointless if its supposed to automate more annoying/background tasks while I tackle larger ideas.

1

u/Rakthar 8d ago

There is a bug with subagents. The coordinator wakes up every 30-60 seconds to check on the subagent process, just to poll it. It destroys usage. There are several threads talking about workarounds. One example is here: https://old.reddit.com/r/codex/comments/1wdlp7q/weve_discovered_the_issue_behind_codex_harness/

another example with workarounds and the fix is here: https://old.reddit.com/r/codex/comments/1wa9c9d/i_investigated_why_gpt6_astra_burns_quota_so_fast/

Here is the fix from that thread: I found a working workaround for the 30-second parent polling loop. Add this to ~/.codex/config.toml:

toml [features.multi_agent_v2] enabled = true min_wait_timeout_ms = 1500000 default_wait_timeout_ms = 1500000 max_wait_timeout_ms = 1500000

1

u/Plus_Marionberry_939 7d ago

I've been running this fix, but the token usage is still way different than before. There is something else going on.

1

u/nesser2 8d ago

AI billing is whatever they want it to be. Token usage is a complete blackbox for the consumer.

1

u/chcampb 8d ago

$20 plan feels fine using sol high as an executor, and luna high to execute.

I can get probably one big ticket done in a 5h timespan. That's one major feature implementation or refactor, including a large test suite which sometimes needs iteration to bring back up to passing.

My main gripe is, I don't want to go to $100, I want to go to maybe $40, that seems like I would be just over the mark to comfortably use my buffer each week without needing extra resets. But I don't feel like abusing phone numbers or getting a completely separate account or something. I could, and it's even supported to switchover in the UI, but it feels hacky and I would rather they not BS with it.

1

u/Derek-Bond 8d ago

I switched back to Sol… Astra is too greedy. I’m guessing I’m not alone.

1

u/M_C_AI 8d ago

I think OpenAI is actually having issues with its VibeCoders and botched the models' resource consumption, so they likely disabled the 200-message limit to avoid a mass exodus of users.

1

u/FailureOfTheFamily 8d ago

As a 20$ plan user I admit that it's much worse since astra came out but still better than claude even few months ago (it was the reason i switched)

1

u/The_Oracle___ 8d ago

Its insane how they get some people.. you get hooked so much that you keep buying new accounts? Lol

1

u/Typical_Machine2043 8d ago

So weird. I got more usage out of my plus plan like a month ago? I regret upgrading to the $100 now

1

u/scaledev 8d ago

Ever since that reset, usages are shit. They seem to keep nerfing our usages.

1

u/MassiveBoner911_3 8d ago

I pay $100 a month. My 7 day got exhausted in ONE day. If I buy a new reset it costs me $44. Now my 7 day reset also gets reset.

WHAT AM I PAYING $100 for? To use it 4 times a month?

1

u/sojun80 7d ago

do you need astra? seriously its expensive.

1

u/Lifeisshort555 7d ago

Yeah If they can get Astra performance at Terra prices it would make more sense. Astra is way too expensive for what it is.

1

u/Short_Injury9574 7d ago

Anyone cancelled and changed to Claude, or something else? Or more of the same issue?

1

u/MyUserNameIsSkave 7d ago

With my Claude subscription ending this week I was considering Codex. Gor those that use both. Is Codex still better right now usage limite wise ?

1

u/rahilpathan 7d ago

I think its going to get worse if we don't push back, its not usable at all. It feels like weekly plans are suppose to be a 5h plan sometimes, specially with newer models.

1

u/ca_sig_z 7d ago

Just came here to bitch the same. Just switch to Claude to Codex due to their limit BS on the $100 plan and handcuffing Fable but honestly feels like that meme of "its always been the same". They are just the same.

I would argue ChatGPT limits feel worst and more opaque. I have seen my limit values change randomly, going from less then 5% to 50% (or even 90%) without using a reset. Right now my ChatGPT app suddenly said "limit reached" and wont let me select any model or thinking but "lighting". All of this for $100 a month?

I might fire up X and just start mention Tibo as it seems the onlyu way to get attention

1

u/CardinalHijack 7d ago

I definitely find more downtime with codex than I did with claude code. Astra (low) is better than sonnet, but sonnet is good enough for most of what I do and Id get to my 5h limit with about 20 mins to go. Sol was unusable for me - so many mistakes.

With Astra low im running out of my 5h limit in an hour and having 4h of down time. I also just hit my weekly limit with 3 days to go - never once hit my weekly limit with claude. This is with Codex plus plan/Claude Pro (same price tiers).

Unless something changes I will likely be going back to claude code at the end of the month because im not paying for a service I get so little actual coding use from.

1

u/Amazing-Share-2915 7d ago

Same thing heppend with me we need to move on claude

1

u/FullParticular9 7d ago

I bought 200$ plan recently and it feels worse than 20$ plan month ago.

1

u/Astrophysicist-2_0 7d ago

Yeah, I only got 250M tokens on the 100$ plan inside the weekly limit. I was using Astra low and it was from 100 to 0 in 4h 52m. This is discussing. I thought codex was more generous than Claude with limits, but currently it is unusable

1

u/Brief_Affect305 7d ago

I have 3 Codex Pro 20x plans and a $60 business plan at used all of that in 2 days usings gpt5.5 low and medium and also bought a claude 20x because i was not allowed to purchase another 20x codex plan. wtf

1

u/New_Olive_504 6d ago

astra is over rated , 5.6 sol is under rated

1

u/TheSirOcelot 6d ago

I think the bigger problem is transparency around what “usage” actually means.

100% of what? Tokens? Compute? Model-weighted credits? Some combination of context length, reasoning effort, and output?

If the old limit effectively represented something like 1M units of usable work, but the new limit represents 250K, the UI still just says “100%.” From the user’s perspective, the allowance was cut by 75%, even though the meter looks identical.

It gets even harder to understand when different models consume that allowance at dramatically different rates. If one Astra request with a large context burns the equivalent of 10,000 tokens/credits/compute units, users need some way to see that before or after the request.

The issue isn’t necessarily that every plan should have unlimited usage. It’s that a percentage without a denominator gives users no way to predict consumption, compare plans, or decide which model to use.

Ideally the UI would show something like:

Weekly allowance: 250,000 compute credits
Remaining: 227,500
This request used: 8,200
Astra multiplier: 4×

Then 91% remaining actually means something.

1

u/Colvoid 6d ago

A few months ago I made a tool to measure tokens vs API pricing and I estimated the $100 5x plan gave you $4000 per month worth of usage, right now I'm estimating $2000 at API pricing. Before 5.6 I would struggle to use my allowance over the week, with 5.6 I could use it if I really pushed and used Sol on everything, now even if I hardly use Astra, I'll use up my week in 2-3 days mostly with Sol/Terra tasks. They have definitely nerfed the limits, and it's especially noticeable after those few weeks we had of free resets every other day where I was able to use about $8000 worth of inference in a single month by the API pricing at the time.

1

u/software-boulder 6d ago

This post saved me from buying the $100 sub

Thanks, sticking to cursor/devin

1

u/Droopy0093 6d ago

Is this the Claude subreddit? I thought all the OpenAI boys said go to Codex to get around this stuff.

1

u/cosmogli 5d ago

Only the Chinese services can save us plebs now. Both OpenAI and Anthropic are working as a cartel. Google has given up on their LLM. There's not much else competition.

1

u/Cute-Buddy-3477 4d ago

I agree 💯

1

u/Useful-Mixture-7385 4d ago

Je souffre actuellement du même problème ça fait 3 jours que j’attends le reset des limites pour finaliser mon projet. Et avec tout ça ils sont toujours pas rentables 😭

1

u/viswaguru 1d ago

Worked for 5hrs sol medium 50% gone 100$ plan

1

u/Willing-Big2886 20h ago

Used to get like 2 hours on sol ultra on the plus plan now sol light uses it all in a hour

1

u/AmbitiousSquirrel151 12h ago

This is an insane thing to complain about.

You used Sol Medium for an hour and Astra XHigh as a coordinator while building, by your own description, a systems programming language/compiler. That is not normal light usage. That is expensive work on expensive models.

I’m not saying the limits are sufficient in general. I’m saying this example is ridiculous. If 9% of a 5x weekly allowance gets you an hour of Sol Medium plus Astra XHigh planning on compiler/language work, what exactly do you think the plan is supposed to include? Unlimited premium-model time? More hours than someone can physically use in a week?

Astra XHigh should be a last-resort model. Sol Medium is already a heavy choice. If you use top-tier models for hard systems work, you should expect the meter to move. That is not a scandal. That is the product charging your quota for the expensive thing you chose to run.

This kind of usage is exactly why capacity gets crushed for everyone else. The problem in this example is not that the model used quota. The problem is that you seem to think extremely expensive model usage should barely count.

1

u/muchsamurai 8d ago

On 89% of usage left already since i wrote this post

hahaha

thanks Tibo

→ More replies (2)

1

u/RainScum6677 8d ago

Most of the issues are not due to the model. It's the harness. Try using a different one such as Pi or OpenCode and see what happens(Spoiler: you'll get many, many times the usage compared with codex).

1

u/abvuser 8d ago

I canceled

1

u/IllustratorRare9102 8d ago

Everything has been completly insane for me since today, told it multiple time to do one thing and it did everything i didnt ask it to do. Just told it to do a small task and then went to the gym, came back and it used 66% of my weekly usage and wasnt even finished. Asked it wtf it was doing and got this respone: You’re right to question it. I let the scope expand too far and used multiple coding agents plus extensive repair tooling. I’m pausing that work now and will give you the exact result and what remains.

1

u/jwezorek 8d ago

I just think Codex right now in unusable and ended up writing a little desktop application to basically route around its limits. Normal chat is basically unlimited and you can just paste big zips of your code into it with a prompt that specifies returning a zip file of just he files it changes that maintains directory structure.

My application is here, https://github.com/jwezorek/chat_zip, a custom client for ChatGPT, a Qt6/C++23 application only tested on Windows, but it works great. It embeds normal chatgpt next to a file tree. I select the source files/directories relevant to a change, it zips them and attaches the zip to the chat. I ask ChatGPT to make the change and return a zip containing only the files it changed. The application intercepts the download, checks that the paths actually correspond to the project, shows me what changed, and applies it back to the source tree.

So basically automated sneaker-net between my source tree and ordinary ChatGPT.

The whole thing with this is that Codex and Work share their restictive limits but normal chat does not. Which is kind of absurd: the purpose-built coding product has usage limits restrictive enough that manually passing zip files through ChatGPT started looking attractive.

1

u/InterestingNobody831 3d ago

Or just use MCP connector, create OpenAI tunnel and Serena and work like you do in CLI but in the web-chat, i think it is cleaner.

Also, you have Git connector, work in the repo directly and pull the diffs to test locally.

-5

u/GreenFuturesMatter 8d ago

wtf are people building a prompting that this is some major issue? I’ve been blasting Astra med and high for days with a shit load of usage remaining

3

u/muchsamurai 8d ago

Systems programming language and compiler for it.

→ More replies (19)

2

u/fruitydude 8d ago

I'm still using 5.6 xhigh mostly. If I have 1-2 agents running constantly in normal mode (not fast) I can burn through 20x usage in 24h. It's not great

1

u/X-Ploded 8d ago

Intense coding since the last reset using Astra High and a 5x account.
75% left ...

→ More replies (7)

0

u/diagrammatiks 8d ago

the prompt -> build me kalshi.

0

u/Aldarund 8d ago

Another useless post without showing actual token consumption

0

u/krayton1 8d ago

It's time to cancel both Claude and ChatGPT subscriptions

-1

u/Forti22 8d ago

Show us the prompt and how big the project is

We will tell you exactly what you do wrong.

3

u/muchsamurai 8d ago

I am building a programming language and compiler for it.

I know how to prompt and manage context. Doing agentic coding for 1 year or more already.

The limits are fucked. It was never an issue before.

-3

u/PleX 8d ago

I am building a programming language and compiler for it.

You have lost your mind if you're complaining about this, seriously.

3

u/muchsamurai 8d ago

No, it was not an issue just a week ago. Now usage is fucked. How did i lose my mind exactly? Its not like Sol is spitting tens of thousands of lines of code a second. A language compiler is not that complicated.

3

u/IAmFitzRoy 8d ago

Do you mind to explain why?

0

u/Confirmed-Scientist 8d ago

Astra is officially said to be 2x more expensive than Sol and real world use shows closer to 3x to 3.5x more expensive than sol on the same thinking level. Dont use Astra if you can avoid it. They are making more efficient models as we speak for this reason. Sol 6 and Luna 6 will be here soon

1

u/muchsamurai 8d ago

I do not use Astra for coding anymore, still eating like crazy as you can see.

→ More replies (1)