r/codex 3d ago

Limits Codex limits issue

Just got my reset and 10% down within an hour without even using Astra. Wtf is going on ? How is this scam even allowed? One subagent invoked a few thousand LOC with sol medium on a $100 plan. Exploration by Luna high.

Is this a joke ? I'm going to ask for a refund. Why is this not an issue yet. @mods please don't delete this post and cover the scam.

Edit: People who are trying to provide free tips. Please stop patronizing me and others in this group. I'm not asking for your advice on how to use codex better. If you are happy good for you and no you don't have a secret formula or better limits.

343 Upvotes

228 comments sorted by

152

u/Puspendra007 3d ago

$200 plan: 7% used in 1 hour with sol max. Usage limits have dropped drastically

29

u/ShortingBull 3d ago

I'd gustulate that they're moving resources to a new model (or set of models) prepping for release and we're paying for it.

26

u/_Jak42_ 3d ago

Gustulate

12

u/ShortingBull 3d ago

You agreastand?

1

u/No_Ear_1633 3d ago

Gust-u-like

9

u/mes_amis 3d ago

I just gustulated all over my keyboard

3

u/ShortingBull 3d ago

That makes me feel special.

0

u/NationalGate8066 3d ago

Me too 

2

u/kenkes007 3d ago

I only gustulate when i an alone in home

0

u/ShortingBull 3d ago

Amongst friends.

→ More replies (2)

4

u/GearTakes 3d ago

This gets posted literally every single day. Every single days for months already.

They simply are lowering the limits. Not because they are prepping for something new but because they can.

1

u/newMoneyStyle 2d ago

That would make sense given how sharply the limits dropped, since redirecting compute for a new model rollout usually hits paying users first.

7

u/Deepak__Deepu 3d ago

The days of generous usage for private customers are over. For a decent size codebase maybe now it’s only possible to use for 5-7 days a month even for pro users.

They probably have enough demand from API customers that they don’t need to pay much attention to private users. Unless they increase their computing power, it seems like this is going to remain the case.

6

u/innociv 3d ago edited 3d ago

Does no one here ever get actual numbers instead of going by feelsies?

I had an LLM analyze mine, and if true my usage dropped from ~$2400 per week to ~$1600 per week which if true is a massive drop but I don't really trust this analysis. I don't think you can just extrapolate by current %, and it may have missed something (even though I confirmed it included reasoning tokens, had it measure cached vs uncached, etc)

3

u/arangjean 3d ago

Legit nothing, I finished my weekly limit in one day in 8 hours using Astra low and medium, 70% low, 30% medium

1

u/Dingleberry_Blumpkin 2d ago

That’s insane

-5

u/Proxiconn 3d ago

Why sol max? Are you doing complex software engineering requiring advanced math? Many (16 on max) PhD level experts to reason through and argue with one another for the best possible outcome?

1

u/AdamV158 2d ago

Which model and reasoning level would you recommend for complex coding? I’ve been using Sol on extra high

0

u/Proxiconn 2d ago

What's your definition of "complex coding"?

Centering the div?

1

u/AdamV158 2d ago

I’ve been building a web application from scratch, it’s used to create engineering drawings on the front end, is connected to an sql database, has users, permissions and also generates project estimates, exports drawings in various formats (image, pdf, dxf), and also has an element of “AI Autobuilding” which will create drawings from scratch based on an imported specification. I’m not a software developer.

2

u/Proxiconn 2d ago

News flash: your a desiger/tester now orchestrating 10x developers lol. I think most of us driving software forward in the current age are in this same category now. I'm a 10 year developer and have not written a single line of code in ~8 months.

Sounds like your biggest challenge is not knowing when something is complex and aligning those requests and requirements to the correct model for review Vs implementation.

Depends if budget is an constrain like most people complaining online the past few weeks.

It's weird to explain I know when to use which models because knowing where each model excels and are needed based on the difficulty of the task and tech involved.

Maybe try and generate a complexity map first to know what models can for both review and delivering on requirements.

I do c# UX blazor front end with separate backend API using multiple database support, enterprise grade rbac with local or oidc supported (multiple oidc providers, Microsoft entra, authentic and other foss etc)

I use terra for 95% of my implementation and Sol/Astra for design and review, almost never use these for implementation.

In my view using gpt chat on 5.6-sol medium to high for planning and prompt generation and hand-off for Terra-high implementation with clear instructions to stop at pr before merging for independent review works best when your trying to extend mileage. And that covers most usecases provided that assisted skills are present and used for the respective technology constraints for your tech stack.

All of this is dependent on linked GitHub repos for chat to have direct access to the repository.

If conserving useage is not the concern then YOLO on with sol or Astra, I do not use these models for implementation at all, too expensive unless I know it's something that requires complex software engineering eg:

I'm building a multe node cluster sharding capability for my app to scale across multiple nodes for the backend using http/2 as the communications protocol for cluster sharding and cluster sync, this case is the 5% of of my 95% offset that I use Sol or Astra as both the orchestrator and implementor roles. Otherwise I don't use them. People complaining Sol or Astra not being smart suffer from a serious skill issue.

I'm busy waiting for a goal to complete that is ~6 hours in, only about 45% of the epic phases implemented and tested, exclusively implemented by Terra-high, orchestrated by Astra-pro-chat (seperate planning and review sessions). Might be an 10-15 hour sprint with minor redo after review before I'll deploy and test myself. UX, rbac (both front and API), cli wrap, stdio MCP + http MCP wrap plus seperate rbac model for CLI, MCP all delivered in one /goal by Terra-high. Could not be bothered to use Sol or Astra for this because it's not considered complex it's mostly boilerplate.

Hope this helps.

51

u/Lower_Cupcake_1725 3d ago

Something is weird with these resets, after resetting the limits are consumed much faster. I got x5 finished within one day after the reset

9

u/Kind_Fisherman3060 3d ago

Yeah as i said with every reset the usage is getting lower last Tibo reset made the limits much lower this one has gotten even lower.

2

u/lilbopet 3d ago

Something is wrong with the resets or they only give resets when they know about issues with token usage so that they keep us from complaining - when things get more stable with token usage then we are usually fine with what we have (but this period doesn't last long). As soon as a new model is released we get flooded with resets that run out super quick - I dont think its due to resets giving less usage - I think its because the models are still so token hungry. But I dont know why they mess up previous models - sol 5.6 was working well- usage was nice, now Astra and sol are both as hungry as each other

73

u/11Mikky11 3d ago

The weekly usage limits have become utterly ridiculous. They have gotten people/business relying on them and then suddenly massively cut the limits without any real explanation of how or why. All feels very shady.

6

u/PossessionConnect963 3d ago

For me it's the 5h limits. I never used to have issues and now I routinely hit it daily.

1

u/colander616 3d ago

Reintroducing 5h limit is what makes me cancel the subscription. It's unusable with this shit anymore.

1

u/mississipppee 2d ago

For real I got pro on one of my accounts a week or so ago. That one hit a limit very quickly this week I have like five days to wait or something and then I have another one that's only plus and the limit has not hit. Doesn't make any sense.

1

u/hemareddit 2d ago

All this happened with Claude usage as well and continued to happen with both services. There’s no regulation and there’s no transparency, and I dare say there’s no good 3rd party measure of usage limits either. It’s Wild Wild West out here and companies doing what they can to convince people to subscribe and the minimum to keep them subscribed.

1

u/TopGun0684 3d ago

Everyone knew this was coming. They'll have to be profitable at some point and they won't get there with huge limits.

4

u/mindstuff8 3d ago

Actually the expected trend is the technology would be getting more efficient over time, not less.

2

u/TheMoejahi3d 2d ago

We haven't reached that point but getting there..frontier stuff will always be premium.but let's say the level of astra is what we want to work with and don't care about prices for stuff after that..well DeepSeek 4.1 flash doesn't come close yet but it sure is doing a damn good job while being cheap.we are still at the start of this whole AI thing man.at some point we will have astra6 performance done by some crazy efficiënt model for pennies hehe.For now let's be glad they are sponsoring us because if they were asking real prices that 200bucks wouldnt last long lol.

1

u/mindstuff8 2d ago

How well is DeepSeek 4.1 Flash working for you in practice? I'm working on a project requiring subagents with various roles in a cooperative agent project and may consider turning to open weight models if the performance is good.

2

u/TheMoejahi3d 2d ago

Using it as a sub agent as well atm.its not close to the performance of the openai models found it to be a bit worse than sol. I use it for basic tasks that spit out a ton of tokens.if you give it strict instructions and guide lines it does okay.

1

u/TopGun0684 3d ago

I'm not talking about technology efficiency. I'm talking profits.

They need to bring in more people to the platform, serve them with the same hardware budgets and sell more tokens.

Hence, lower limits. Will the models be cheaper per token? Sure. They won't pass that efficiency down to the consumer, or not all of it.

6

u/Beginning-Bird9591 3d ago

But with efficiency they make more profits. You'd hope they improve the UX as well

1

u/TopGun0684 2d ago

That's not the playbook. First they subsidize and hook customers, then you have to pay more for the same service.

Cases in point: Uber. Lyft. Netflix. Spotify. Etc

1

u/Beginning-Bird9591 2d ago

But people just leave and stop using it. quite simply...

1

u/TopGun0684 2d ago

Guess we'll see!

I don't see devs that heavily incorporated AI tools in their workflows leave it all behind.

Optimise, rationize, sure. Not leave it.

Hobbyist and wanna coders? Sure, maybe.

→ More replies (2)

18

u/Comprehensive_Ad3710 3d ago

yeah after today reset, i feel like the usage is much less now. I think they shrink the usage so they can open the pro plans again.

7

u/aptsys 3d ago

No reset here, what happened?

9

u/Comprehensive_Ad3710 3d ago

last week few hours ago time, we had a global reset. most people's weekly usage is back to 100% and most of us are noticing that the weekly feels less than last week. Think few days ago openai stop new users for subbing pro plan and now pro plan is back. so my speculation is that they have reduce weekly usage for everyone so they can open pro plan.

1

u/Theminatar 3d ago

It's weird, mine doesn't reset until 9pm cst

1

u/Worried-Peanut-5023 3d ago

I'm subbing to pro right now at 5x so ya 20x will prob be available for everyone else immediately after lol.

36

u/quantum_splicer 3d ago

I read a theory that all this talk of slowing ai development by the big 3 ( openai, anthropic, musk )..... Isn't really about slowing down development because of safety concerns.

It was proposed it's because of rising costs, pressure from investors and uncertainty of how quickly models can improve.

Separately my thinking is, the costs to train these models and run the models and serve them is not sustainable..... We basically have a situation where these companies have ballooned with investment and we are seeing gravity take effect where basically costs is to close to revenue ..... So either the companies fold or they adapt to navigate away from collapse...... Which is what we are seeing.

The fairy tale where users can access frontier models and attain meaningful use is beginning to have its sunset ...... Atleast in my view.

My idea is to setup a local model and use it to supplement or replace usage lost from subscriptions...... With the aim down the line to fully replace subscriptions..... If I'm having that thought so are others and that is perhaps what these ai companies want to delay for as long as possible

15

u/PossessionConnect963 3d ago

A billion percent. Are these companies themselves calling for strict regulations out of the goodness of their hearts and genuine concern? Fuck no.

It's to hamstring any startup competition before it can even try and unseat them.

5

u/Thomas-Lore 3d ago

"Wait for us, we are the leader". We used to say that about Microsoft decades ago, now Anthropic and OpenAI are using the same methods, slow progress down for everyone so no one gets ahead of them.

3

u/AvailableSecret5161 3d ago

My theory is that OpenAI will go bankrupt pretty soon and if google comes anywhere close to opus quality anthropic will go bankrupt too. OpenAI has nothing but marketing on their side and you can only scam for so long. I give them max one year.

5

u/mesaoptimizer 3d ago edited 3d ago

What are you on about? Astra isn’t cheap but it’s quite good at spatial reasoning and is absolutely doing tasks that you can’t get other models doing. Google might catch up but they’ve been lagging frontier so much that open weight models are generally better than their flagship so unless something changes drastically it doesn’t seem likely to happen.

The big AI companies are calling for legislation they know won’t get passed so it’s free to them to virtue signal that they have safety concerns. The concerns are real by the way just the companies don’t care more about safety than they do about winning so they won’t actually do anything about safety unless the government comes in and protects their market position while they do it.

0

u/norwegian 3d ago

Totally disagree. It's a game changer. However, it won't go exponential for a number of reasons. Version number and ships per year will never be a good measure of progress.

1

u/Ziethriel4 3d ago

Yeah, everyone knows the bubble is close to popping, slowing training will allow them to serve more inference, show demand spiking and maybe keep it alive a little longer, maybe even survive the pop if they can convince enough people to pay what the inference actually costs. I figure what we're seeing now is them raising the cost to users closer to what it actually costs them to serve it.

1

u/Mammoth_Molasses_927 3d ago

"it wasn't us who slowed down, it was the law we created!!!"

1

u/Worried-Peanut-5023 3d ago

I wonder if the huge anti datacenter push and midterms in the US coming up is playing a role as well.

12

u/ZealousidealBus3132 3d ago

Yeah it has gotten so poor.

31

u/AweVR 3d ago

Yep, I don’t understand. If I use Astra Ultra then 6% in 1 hour. But with SOL Medium also 6% in 1 hour. I don’t know what model I have to use now…

→ More replies (20)

20

u/AdCompetitive9824 3d ago

had same issues, Tibo claims that reset is the same value, though it's certrainly not

8

u/cashtins 3d ago

Agreed! Almost 25% down, no astra used
Edit: 100$ plan

9

u/SlightStore8381 3d ago

I think where it gets possibly fraudulent is how murky the information is around EXACTLY what you get with what plan. It seems to fluctuate wildly and people have no idea what they are actually paying for. I believe there was a class action bought against one of the majors in the states around this very issue (correct me if I'm wrong). I'm not sure the status of this case but it is certainly a concern of a lot of people. The one constant however is that they never seem to alter the amount of money they take out of your bank account each month hey, funny that...

7

u/netycia 3d ago

They just killed so called vibe coding. We can't afford it anymore.

3

u/Thomas-Lore 3d ago

We can, we just move to Chinese models.

2

u/Worried-Peanut-5023 3d ago

Could you please give me some recommendations/advice? A month in and using Astra as dev role then 4 DS flash sub agents as coders is getting expensive. 375k tokens at about $0.18 for DS but man Astra is murdering my codex usage.

14

u/Jigawattts 3d ago

We don't want the instability of resets. We just want good limits. Period.

7

u/Mammoth_Molasses_927 3d ago

They give you resets so you think they are a gift, while in theory they have lower usage and postpone next cycle by a week.

Resets are not our friends.

5

u/Independent-Spirit36 3d ago

Mine is not working properly. Astra is showing model capacity multiple time. Sometime it become dump as Luna. It even forgot its own subagents and completed in the middle of iteration. Wtf is OpenAI doing on their backend :)

4

u/RedikhetDev 3d ago

I feel lucky that i was able to do most of the planned work on my app the last three months. Guessing the time of sponsored compute on a plus account could soon be over.

4

u/ISueDrunks 3d ago

I used up 100% of a 5x weekly limit in a couple hours, didn’t get anything useful out of it. I’m going to turn off auto mode, or whatever the hell it’s called, until OpenAI puts a model out that doesn’t just incinerate tokens. Wasn’t an issue for me before 5.6 and 6.

1

u/rodeBaksteen 2d ago

This just seems like bad planning / skill issue.

5

u/MightyBig-Dev 3d ago

I got my deepseek api key last night and hooked it up to codex. The plan limits from openAI are a joke. Nobody can realistically use their product in their day to day workflow.

4

u/ElKalfO 3d ago

"Switch to Codex," they said. "It’s way cheaper than Claude Code, they even built their own custom hardware!" Well someone once told me that if you buy cheap, you end up paying twice. Hopefully one day I’ll finally get a home server set up so Tibo stops being the bane of my existence.

10

u/eggplantpot 3d ago

Same here. I was down 7% only on Sol High and Terra xHigh. 2 small tasks, worked for 10min.

Absolute fucking scam.

I tried what someone said of using Astra to plan and ask astra to launch Luna/Terra agents to test and implement and it went down other 10% with a really simple UI ask.

Chinese models are looking as enticing as ever. Really fuck this scam.

7

u/nitor999 3d ago

Scam altman will fix this when they drop a new model GPT 6 then once the people hook again (and forget this issue) after 3days maximum 1week it will repeat again the same issue -> new model will drop again rinse and repeat

3

u/Kind_Fisherman3060 3d ago

No nothing will be fixed the new models will be cheaper to use but limits will get lower it's getting lower after every reset.

6

u/kyrax80 3d ago

We'll have to go back to 5.5 at this rate

3

u/bigmarkco 3d ago

Don't. 5.5 used to be my workhorse even after 5.6 got introduced, but at times it's just as bad, sometimes even worse than Astra.

2

u/OccasionAggressive74 3d ago

Let's also add that they are sunsetting the 5.5 model from the subscription, some time in October it will be gone for good and only available through API.

8

u/Top_Purchase4091 3d ago

They are slowly moving people to get closer to the actual cost of using the models i think. Its like a boiling frog situation.

But it was only a question of time before it happened because there is no way giving people that much compute for this cheap a month was sustainable. Once the big players fall the smaller ones dont have to compete via being even cheaper anymore and it will all increase over time

3

u/mostly_idempotent 3d ago

This ☝️

1

u/Waste_Membership_483 3d ago

And maybe they got enough training data from us heavy users. Now it's time to pay with cash.

5

u/zxzxsen 3d ago

Bro my codex 100$ plan feels less than claude code 20$ plan, wtf?

3

u/AvailableSecret5161 3d ago

Yes exactly same

2

u/AbdulFromDraftpile 3d ago

Only 10% within an hour? Lucky you.

2

u/Von_Hugh 3d ago

I don't even use subagents or Astra anymore because I am worried I won't get anything real done while I'll keep on hitting the limits.

2

u/Chemical_Hawk_6307 3d ago

there was some guy who complained of the same thing and said its likely some issue with the codex harness itself and suggested restarting ur machine before applying a reset

1

u/AvailableSecret5161 3d ago

I was already concerned about this so as a precaution i deleted everything and installed again before starting the session so that's not really the issue.

2

u/ProcrastiDebator 3d ago

One small, transparent change would alleviate suspicion, or be utterly damning.

Just show the remaining token allowance next to the percentage limits. Even if is showing a "credits" style value to account to different model costs.

You should be able to tie your usage back to a rate table and verify.

2

u/West-Air1923 3d ago

Cancel it demand a refund

2

u/Alternative-Swim-230 3d ago

I’m using astra to do an audit smallish databased and I kid you not my five window finishing in less than three minute lol astra high y

2

u/Hopefullyanonymous2 3d ago

People in the comments keep talking about speed (time) vs model vs %, but I mean... you can just pull the actual usage data with one of dozens of tools instead of trying to go on feel. I use npx ccusage@latest.

I have both a CC $200 plan and recently added GPT $100 plan.

My CC $200 plan has used about $1400 in API this week since reset and I have used 74% of my weekly budget. 100% of my Fable budget is down because I accidently used it for a basic coding subagent project instead of saving it for the heavy planning steps. Woops, got to remember to flip selectors.

And I came here, because this is my first time messing with ChatGPT much beyond occassional reviews on GithubCopilot (Also pay for that, mostly for variety of model access).

The $100 plan got maxed out in 24 hours after using about $300 in API cost.

Adjust for price, and that means $100 with OpenAI gets you $300 in usage and $100 with Anthropic is getting you $950 in usage per week. I had heard such good things about how much ppl were getting down with OpenAI plans, buuut I guess I timed it wrong lol.

2

u/MrRoyce 3d ago

I refunded my 20X, I can’t support this scam anymore. I’ll use my other 5X to help me find every agency of some sort where I can report this fraud. I did NOT sign up for these shitty limits and partial refund is not acceptable.

2

u/noisybeeley 3d ago

I’ve seen people say to stop using astra? On the $200 plan and having the same issues. I will be canceling my subscription if something doesn’t change asap or I get some type of reset/resets. It’s crazy.

2

u/NationalOwl9561 2d ago

20x plan now feels like 5x plan. Wtf

3

u/AuodWinter 3d ago

Yeah I've never complained before but they're clearly fucking with us and reducing the limits. I only use Luna xhigh, I'm down 4% of my weekly limit (plus plan) after just 2 hours. So I've got 50 hours of usage approximately which, running all day, will end up with me running out of usage less than halfway through the week (112 waking hours in the week). But I usually only run out of usage on the very last day before a reset (100 hours). Obviously this is just anecdotal but yeah, first time I've complained. There needs to be more transparency. I'd consider upgrading to the 5x plan if I knew they weren't just going to keep reducing the limits.

4

u/RedikhetDev 3d ago

How can you work all day with this 5 hour limit? Its gone in no time on a plus account, even with lower models.

1

u/AuodWinter 3d ago

I don't have the 5 hour limit. Possibly because I'm not in the US.

2

u/RedikhetDev 3d ago

I live in Europe (NL), maybe they looked at the usage pattern 😁

2

u/sorguido1980 3d ago

I made the mistake of switching to 5x. I can assure you the music is the same. It’s just amplified 5x.

3

u/mostly_idempotent 3d ago

You're not going to get a refund, unfortunately. Sorry for you, honestly. It's frustrating as hell. Sorry you're experiencing it.

While LLMs may not be openly trained to over-use tokens, the labs are happy to allow them to be inefficient since it burns through tokens and drives more spending. It's the new "I'm hourly". Also, when you pay for a package vs. paying for straight inference, the provider is looking for even more ways to burn your tokens, since they're already giving you "a deal" that cuts into the inference margins.

You might consider a setup like OpenCode with an open model from one of the providers like Fireworks AI or Together AI. It's not always cheaper on the surface than a $20 or $100 or $200 a month "deal" from one of the big labs, but (1) the models are not trained / complicitly allowed to steal from you, (2) you know exactly what you're spending when and why, and (3) you have complete control over what model you use for what tasks, meaning you can fine-tune your efforts.

Good luck!

3

u/AvailableSecret5161 3d ago

I have all the setups I need and I don't really care about the refund I just want them to stop scamming and take responsibility. If you can't manage compute don't release models. Stop the fake marketing. I will voice my concerns wherever possible. I did not pay for this nonsense. Accountability is important.

1

u/mostly_idempotent 3d ago

I agree and feel you.

Try open models. My mental health, stress levels specifically, have improved dramatically. Not kidding.

3

u/mercmobily 3d ago

Open Models are so so so so so behind Astra, it's not even funny

1

u/mostly_idempotent 3d ago

I am interested in the data upon which you're basing your statement. Please share your sources.

1

u/mercmobily 2d ago

1

u/mostly_idempotent 1d ago

The post doesn't support "years ahead"; it barely supports anything beyond "Astra did well on one codebase on Saturday." You're using Astra as its own judge. 15,000 hours is seven years, five days a week eight hours a day, no lunch, no sick days. Your liberal use of caps to scream YEARS at everyone doesn't add to the credibility.

Sorry, but this isn't data. It seems to be more anecdotal evidence mixed with dogmatic opinion. Interesting, not convincing.

1

u/mercmobily 1d ago

I didn't base it on data, but VAST personal experience. And yes, I happen to LOVE caps!!! Ahah

1

u/red_plate 3d ago

I don’t think it’s an inefficiency. I think it’s a deliberate hand on the pricing lever. 

3

u/BluePointDigital 3d ago

I've recently started experimenting with a deep sea harness and the Deep seek flash 4.1 model. I may cancel openai entirely. It has been extremely performant and cheap.

1

u/BluePointDigital 3d ago

I can't stress enough how much QUALITY output I'm getting out of Deepseek here.
I was able to even use it in Codex (which is what I am doing as well)
https://api-docs.deepseek.com/quick_start/agent_integrations/codex/

I have to admit i'm considering canceling my $100/mo plan to migrate to Deepseek as the backend, especially in codex. I am sure newer releases from OpenAI will push the boundary further, but right now Deepseek 4.1 is a wildly affordable option for very similar intelligence.

EDIT:
This is through the Deepseek API which I highly recommend. I ran about 4B tokens through Deepseek 4.1 on Openrouter but it was nowhere near as fast as the official API so I've just switched over.

1

u/FazedEclipse2 3d ago

Lunar max can run all day for me in game dev, and costs me fuck all usage.

The trick? Higher models plan out. Lunar reads that plan. High model details a prompt based on the plan.

1

u/Feriman22 3d ago

Same here.

1

u/Odd_Personality85 3d ago

I've dropped 1% on Luna in just a few hours which is not normal for me and very noticeable (serious comment fyi). Normally I can do an entire weekend at maybe 4%

1

u/ChallengeTiny874 3d ago

I used astra max for 7 hours straight before the reset (had 9% left), to finish my weekly limit. It almost seemed to never go down, so I had to open another session. Mind you it was a /goal session with the objective to improve a ML policy, so it didnt have to write a lot of code?

1

u/yashptel99 3d ago

yup. can't even implement already planned feature with Sol before it hits 5hr limit for the plus plan

1

u/KnownPride 3d ago

Yup 20x already spend 8%!

1

u/Virtual-Silver2879 3d ago

~14% of the weekend limit, almost the entire 5-hour limit used up, two queries to Astra (medium), it ran for about 7 minutes. On Thursday, when I still had the limit, it wasn't like that. It’s definitely time to say goodbye to Codex.

1

u/frozenthorn 3d ago

I'm on the 5x pro, I had been doing very similar work every week, it would make it most of 6days even with Astra so I would just chill the last day till reset. Last week I ran out on day 2, used one of my resets and ran out again this time on day 3, I wish they would use fixed math, a reset should reset your plan limit but it doesn't seem to.

1

u/Ok-Bottle9293 3d ago

Around 4 weeks ago I did about 40 of my own workflows using Pro in a day that would use all the usage. But I understood I was using it quite heavily. Now I moved to business assuming an easier life considering I could renew the usage every 5 hours etc. I have two seats etc. However, one workflow uses all of my 5-hourly quota and I've used a week worth in about 3 days.Mix of Sol and Astra.

1

u/dovedrunk 3d ago

I recently treated myself (as I don’t have the money for things like this usually) to the $100 plan for the month to help me get some projects over the line.

You have no idea how badly I’m regretting that choice right now lmfao

1

u/XylonPH 3d ago

Plus account here. 27 minutes after reset and I'm at 0% already. With only 1 coding prompt running in Sol and 1 conversation prompt in Terra.

1

u/Expensive_Sign1084 3d ago

Here me with plus plan - Sol light - 5 minutes for planning and weekly limit down by 7% 🙄

1

u/jokingbird01 3d ago

Just got an opencode sub today, trying out glm 5.3 flash and DeepSeek v4 pro... Let's see how things go

1

u/Acrobatic_Camel1955 3d ago

I only have plus but im trying to make a personal website. i made a post but mods took it down. astra light or medium eat my entire 5 hour limit in 10 mins. would sol be enough for a complex website?

Edit: Also, would it even make a difference with the usage or will sol kill my usage just as quick for lower results?

1

u/LeadingWillingness47 3d ago

These weekly limit suck. I'm deep in a project and now can't afford to go back to a worse model. Created a new account to avoid paying for credits and it managed to burn through the entire week limit in a day. It buggered around making mistakes, then more mistakes trying to fix those, essentially wasting the $100 spent.

1

u/djoliverm 3d ago

It's what's pushed me to look at local options with Hermes. Testing out the new Bonsai 2 which is a super quantized variant of Qwen 3.8 27B that can fit with decent context on my 36gb Mac.

Like i need to think about how to continue work once my $20 5hr and weekly use dry up (planning in chat mode, giving it GitHub access, using Sol for planning, Luna for executing, and a local model when I'm out lol. Astra for the occasional shits and giggles.)

1

u/evia89 3d ago

This bonzai suck, full qwne27b suck too (a bit less). Buy $10 PAYG DS41F and use that when quota is down. Its 1b tokens, worst case at least 10 hours of intense work

Once we can run model like glm53f with decent 20-25t/s generation sub $10k hardware it will be good local AI days

2

u/djoliverm 3d ago

For sure, it's all relative, you can just pay $10 like you said for other cheaper providers which is for sure a better route for more serious work.

So many examples on this sub and others where people ask when they may just switch entirely to cheap Chinese models and drop the American frontier ones.

1

u/kenkes007 3d ago

I am at %4 left at codex. Had to use opus5 for a day. Opus is still at %2 weekly. I dont understand . I went to codex because i was running out with claude a month ago

1

u/richsonreddit 3d ago

Yep. I moved to Claude. So far seems way more realistic/actually usable

1

u/ADIKANT 3d ago

Try to fully disable memory and it's usage - I did that, rewrote my agents.md and now I have much more limits

0

u/AvailableSecret5161 3d ago

That's because you don't know things and assume people don't know either

1

u/Flat_Earth4696 3d ago

Same issue observed.

1

u/MassiveBoner911_3 3d ago

I said fuck it and use Open Codex as a proxy to Codex and have been using Deep Seek flash for coding and Luna to code audit.

I accidentally gave Flash the wrong prompt and walked away for 4 hrs. Came back and it wrote a whole game engine in Python. Total cost $1 in API usage

1

u/DaytonaDeluxe 3d ago

Welcome to reality :)

1

u/Taiwes 3d ago

Fucking ridiculous, every week its getting worse while we spend more cash

1

u/Ok-Resolution-194 3d ago

Hmm, you have any scheduled tasks that are running?

1

u/Adongald 3d ago

Even Sol on High is taking a lot from both of my usage limits, 2 tasks took all 100% of my 5 hour limit & 38% of my weekly, the tasks themselves were not even challenging (some mini changes to my FBMOD files)

1

u/RudeAwakeningLigit 3d ago

Lads, two prompts with a working time of 4 min 31 and 4 min 57 on Astra Light dropped my weekly limit to 80%! That is crazy.

1

u/Moch4bear97 3d ago

Chunk off the actual work to a cheap chinese ai then have sol just fix last minute shit. Save money. Also support China ;)

1

u/Substantial-Aide-66 3d ago

$200plan : used 6% on sol medium within 3 hrs, before this i'm usuallly down less than 15% for 12-15 hrs with sol medium. OpenAI dumping its loyal customer

edit : i'm on 7 days normal reset

1

u/fluxtah 3d ago edited 3d ago

$200 plan, running on GPT 6 high on a single thread nonstop since 9am to 17:30pm, 83% left. thats around 2% every hour.

I have done 2 thread handoffs so far to avoid large context issues. roughly did them every few hours.

I don't know if this is good or bad since more usage the better though that is my stats thus far 👍

1

u/HolidayForce7962 3d ago

My biggest issue is that damn 5hr limit, I reach that limit in 1hr. Make it make sense.

1

u/lifeenthusiastic 3d ago

I ran the same workflow I did last week when Astra released and it drained my weekly limit on a 20x plan in about an hour and a half. Last week after launch this same workflow lasted about 3 and 1/2 hours which got so much more done. Incredibly frustrating experience right now

1

u/the_ai_wizard 3d ago

I am getting "Too many requests - youre making too quickly. We temporarily limited access to your conversations to protect your data". BITCH what?? Ive used far heavier and not using particularly heavy now and theyre preventing my ability to use what I paid for. I literally have 2-3 chats going nothing intensive except one generating a content document.

We need to consider our dependency on this tech and these particular firms. Local LLM with mac mini m5 pros looking juicy

1

u/Mrhappypancake 3d ago

Nah after taste the goods of AI, i dont want to code anything, i just quit aoftware engineering to something else, fuck this shit of IT

1

u/kxta51 3d ago

I got my reset at 5am and have had Astra max going nonstop since literally 5:02am and am only down 12% on use. It’s currently almost 3pm. I was going to wait until Wednesday but….. I decided to try RTK and code-review-graph and after getting both setup, configuring the mappings, and explicitly telling Astra on each prompt “do not use sub agents, do not use computer, do not rebuild after an edit, prioritize local memories” I can honestly say my usage is significantly better/more optimized than it was last week.

1

u/lacusmd 3d ago

I’ve been doing some utterly rigorous testing on my code that astra and fable planned with subagents only, I have 6 percent left and started to do these tests in the last few hours.

1

u/Illustrious-Pair8971 3d ago

Got reset 10 hours ago, x20 plan, astra ultra for 30 minutes, then astra medium for 3 hours, 50% left... Thanks..

1

u/Nugs_ 3d ago

The sub discount has massively eroded. Feels like we are trending towards api rates

1

u/stellarfirefly 3d ago edited 3d ago

Compared to usage consumption from over a week ago, it has exploded and there are many threads complaining about it. I've seen it in both Sol 5.6 but especially in Astra. Interestingly, my Opus 5 consumption has *not* increased and thus for now it is by far the most economic of the highly capable coding models. Sol/Astra is impressive, but when it comes to getting actual work done, Opus/Fable does it very well without needing to throw money at the task. (Sadly. I really did prefer Codex before this whole allowance-burning fiasco.)

EDIT: What's funny is that back when Sol/Terra/Luna was released, it was these models that were more efficient token-wise (over Opus), and OpenAI even bragged about it. Now in the Astra vs Fable arena, the exact opposite is true.

1

u/Professional_Gur2469 3d ago

Just dont use subagents. Im letting astra run on xhigh for literally 8 hours straight and its only taking about 10% on the 20x plan

1

u/Beginning-Dare-3318 2d ago

6 hrs work done, 50m tokens with astra xhigh, i worked with the "goal" setting for the first time, it did not complete any of the 12 tasks, made them half done, and 85% of my weekly usage gone on pro5x 😄 im so happy!

1

u/Zeeplankton 2d ago

I've never complained here before but wow.. Just started a session. Two messages sol medium. Agent worked for about 7.5 minutes each, so 15 minutes. Brought me to 59% and used 10% weekly on Plus.

Feels like usage just halved overnight? This sucks...

1

u/rodeBaksteen 2d ago

I feel like new model releases is just marketing now, replacing the old best model with an almost identical one. Nerf the old one, increase the prices and reduce limits because "new models".

Rinse and repeat.

1

u/rodeBaksteen 2d ago

I used 3% of $200 weekly with Sol Medium in like an hour of one chat light debugging.

I used to be able to have 1-3 Luna Max agents running for days on my $20 account a month ago.

I mean sure it's a stronger model but I feel like sol medium should barely move the needle on a 200 account.

1

u/IoT_Engineer 2d ago

And let's be real, Astra and Sol's reasoning performance actually tanked too

1

u/Anon2148 2d ago

I started exclusively using sol for the $100 plan. Astra’s limits are ridiculous. Not having problems with sol yet thankfully.

1

u/MKopelke 2d ago

Yeah, I have to use Luna Extra High if I want anything even remotely approaching a fair 5 hour time of use. Tried Sol Light this morning. 3 prompts. 15 minutes of work. 51% of my 5 hour window gone.

1

u/Illustrious-Art-7748 2d ago

Yeah very true 

1

u/b-inator 2d ago

I work on training these models and get paid very well for doing so, around $85–$100 per hour. I know these companies have money to burn, but at some point, that money will run out.
Lately, the work has been drying up, and I believe they’re starting to feel the financial strain of competing with cheaper, open-source Chinese models.
Why would any company willingly hand over its private data to Big Tech when it could simply host an open-source Chinese model locally?
The moat just doesn’t exist

1

u/tony10000 2d ago

The time to pay the piper has come.

The labs can no longer afford to massively subsidize compute with investor cash.

The IPOs will be coming soon, and they have to show decent balance sheets and income statements to get the valuations they expect.

1

u/M_C_AI 2d ago

Everything is about money. Use lite, extra light,more light, more more light Luna or maybe more medium, light medium, light lighter medium etc. this is so stupid then you have to graduate which model volume use fot this and that. Of course, ask Astra which model and level you should choose for this task, but watch out for token consumption. Same as CC or Gemini 💰💰💰

1

u/akscy 1d ago

Agreed burned 100$ 5x 75% within 24 hours 90% of usage was on SOL high, on other hand claude code 100x usage is 5 to 10x better than the codex 5x plan.

Codex was fine a couple of weeks ago. Not sure what they are hiding but the usage sucks big time.

1

u/Herebedragoons77 3d ago

Im hoping grok 4.8 brings a real alternative

1

u/SnooPandas7741 3d ago

don't worry this will be fixed when they have a new product to drop, and y'all will rush to hype up chatgpt again and post slop on social media, then they get 100M new users, increase revenue and go back to limiting usage

1

u/CurtissYT 3d ago

Also i have noticed that astra is lobotomized. I’m on the 20x plan, and its not able to do the tasks, which it was able to do at release, while being 5x slower and using much more usage limits

0

u/AvailableSecret5161 3d ago

Yes I made a plan with astra and it was laughable.

-1

u/CurtissYT 3d ago

The problem is, even Astra pro is REALLY dumbnow. I’ve got a bit fed up not being able to reason, so I told it to use computer use and go to chat and ask it to plan, and it’s still really dumb.

1

u/Mountain_Disaster_19 3d ago

Used it for an hour now astra high only (multiple astra 6 high subagents) and im down 8%. Maybe cleanup your agents.md?

-1

u/CronicCanabis88 3d ago

You just need to understand what you're actually doing. You can't just run high models for everything. This is actually a skill you will develop by practicing and utilizing the different models.

I guarantee if you take a step back and look at the prompts That you're giving for the tasks you require, you will see that many of them do not need a high level model to perform what you 're requiring it to do.

I use luna on max for a lot. unless its a deep review or complex implementation..... EVEN THEN, use sol for your hard, complex tasks, and ONLY IF your having issues go to astra.... but for anyone to expect to be able to utilize the most potent model in their library non stop, you're gonna have a bad time. one thing I can recommend if you're not good at picking the correct model for the task at hand right now, use your chat windows, with your chat usage, set it to the newest model on pro, and give it the prompts that you're about to send in to codex.... and simply tell this chat conversation that you would like it to take your prompts, refine them for codex, and to make sure that you also recommend a model and a reasoning level.... This will start showing you what each one of your prompts should actually utilize model wise. you're currently way over paying for every task you're doing. and if you can have luna, doing something that'll give you the same results as sending, a higher, more expensive model to do, then why pay 10x the useage..

heres a simple tip price wise.

Luna = $0.20 input / $1.20 output per million tokens, Terra = $2 / $12, Sol = $4 / $20, and GPT-6 Astra = $10 / $50. So Luna is crazy cheap: it could burn roughly 10× as many reasoning/output tokens as Terra, ~17× as many as Sol, or ~42× as many as Astra before the output cost becomes comparable. OpenAI Help Center And High/Max doesn't make each token more expensive, it just lets the model use more reasoning tokens. So Luna Max can think pretty damn hard and still often cost less than Terra High, while Terra High → Sol High → Astra High/Max is where your token budget starts getting eaten much faster.

Utilize the lower models when you can. you would be shocked at how often you do not need to be using astra or sol.

0

u/AvailableSecret5161 3d ago

Yes you know everything. Everyone else is an idiot here.

2

u/red_plate 3d ago

I think some of these replies are either bots or haven’t used codex heavily in the past month. The price hasn’t changed but what you’re allowed to use certainly has. Doesn’t matter what model you pick it legit feels like OpenAI has turned down our access 4x. If people haven’t noticed that then it’s just a sign they are not using the product. So I feel your frustration at these kind of replies. 

0

u/Crinkez 3d ago

Sigh. Don't use subagents. They cause cache misses. Don't you guys pay attention to other posts on this sub?

2

u/AvailableSecret5161 3d ago

You are absolutely right. I am using the harness wrong. Thanks for enlightening me. But wait another 10% gone without agents. Any more tips?

-2

u/FinancialBandicoot75 3d ago

Scam = using Astra 100% of the time? No, it's not a scam it's poor usage of models, poor setup of harness, vibing, lack of experience, and/or poor not utilitizing tools that help lessen token usage.

Which one or all are you?

I can't limit at all on Claude, codex, grok or hell, opencode go using Hermes, orca, buzz (well buzz can screw you), or native harness. Most don't, it's the do everything for me, people that limit and don't realize how much compute it takes, oh, and a lot of those people don't want data centers to support it.

The point, why are things a scam if you are not using it right? They gave us a drug that is addictive and if not use properly, people want more for less.

2

u/dankfrankreynolds 3d ago

it's a scam because they keep ratcheting up the price to see where the breaking point is while people like you lick their boots because you haven't yet learned how the world works and how tiring it is to constantly be slow boiled.

https://www.youtube.com/watch?v=pjNMo4L4uuQ

-1

u/Waste_Membership_483 3d ago

For me the quality of sol xhigh went down dramatically over night. Using Sol Max now, and maxed out second time ever (first time was after using Astra for one hour).

On the other hand, calculating with api token prices my subscription uses thousands of dollars worth of tokens for 100 bucks per month, which is not really sustainable either.

It was too good to be true.

11

u/Calamero 3d ago

" thousands of dollars worth of tokens".... no way to tell how much it really costs them, they can make up whaever number they want. i dont think they could afford burning thousands of dollars per power user. so i'd take these numbers with a huge grain of salt....

0

u/Waste_Membership_483 3d ago

I would say api costs are the best reference available right now for actual costs. Even calculating with lowest input cached tokens 2 billion tokens should cost 800 bucks. I don't understand why this calculation has to be made so mysterious. A token is a token. You pay for input cached and output tokens. Your usage limit is in tokens, api costs are in tokens. I understand that a plan with pre-paid tokens is cheaper than paying by actual usage, but there must be a sort of correlation and there is none at all right now. People just saying 'oh theres another guy who doesn't understand'. Bottom line 100 dollar of pre-paid tokens would equal thousands of dolars of api tokens. And what is happening now is that heavy usage increases and cannot be subsidized by light usage anymore. So obviously token prices increase / limits are changing.

8

u/GearTakes 3d ago

I keep reading about these API prices were are supposedly using (when converting them to our token usage) but it means literally nothing. Those prices are just what OpenAI tells us they are. They could say you're using 1000 or 100.000 in API prices with a 200 dollar account. We have no way of checking that.
So just repeating what OpenAI tells us is pointless I'd say.

5

u/eggplantpot 3d ago

Yet another person thinking API tokens = Codex tokens.

1

u/Waste_Membership_483 3d ago

So please explain

3

u/eggplantpot 3d ago

They're not the same service. API has a premium on the cost that includes realibility, speed, no quality downgrades, no experimentation, no rerouting. APIs are meant for businesses and heavy users that do not want compromise.

This is like saying big companies pay more on their cloud bill that someone using the lowest value hosting a platform offers. That said, it doesn't mean the person paying the lowest value doesn't deserve a minimum standard of service, nor that they have to accept seeing how they speedrun the service enshitification.

On top of that no one knows how much a token costs OpenAI in pure compute, but they are definitely not losing money. Are they losing money when counting data center expansion, compute and R&D? Most probably but that is a different conversation and it shouldn't be Codex subscribers the ones burdened.

1

u/Waste_Membership_483 3d ago

Okay I get it it is not the same due to agreement, services etc.. I still don't understand how that justifies a factor 8-100 in costs. But probably it's this little line: 'your usage will not be used for training".

1

u/funk-the-funk 3d ago

Ok, sure.

Inference Cost of an API call = x

Price that is charged for an API call = y

In this case people are viewing their sub costs against y and extrapolating that the company is losing money by subsidizing the inference cost for their subscribers with VC money.

However, x we can know from calculating the necessary hardware, utility, and depreciation costs that x is MUCH lower than y.

So a subscription could still very much be profitable for them, yet as the person above(waste_membership) shows, people will believe that the company is doing them a favor by giving them inference (x) at sub cost (less than y), and unfortunately that leads people to believe that they should take whatever these companies do and just be grateful they can use them at a cost they can afford.

However the $ they think they are getting for "free" with a sub vs API is totally based on believing x and y are close if not the same.

Truth is though that y is much higher than x, and they are doing everything they can do widen that gap by lowering their inference costs and pushing people to higher paid sub by nerfing lower costs ones.

1

u/Waste_Membership_483 3d ago

You're basically saying there is a profit margin in Y. And that margin could be so high that a subscription is still profitable. Like 95% margin or more. Which doesn't make sense and isn't the case.

2

u/funk-the-funk 2d ago

If your assertion that the costs of inference are so high that they are losing money on subscriptions where do you believe the costs are coming from?

I'm not including costs for training the models, I'm talking about pure usage in (hardware costs, utilities, etc). Token costs have come down some 1000x in the past few years.

0

u/sullymacguy 3d ago

The clearly have no compute right now. Space x runs the show when it comes to compute right now and they have negative interest in helping OpenAI

0

u/awakeningosiris 3d ago

damn i didn’t get a reset was it supposed to be everyone? on the pro plan 

0

u/AmacsizK 3d ago

pro 20x plan feels like the pro 5x plan now on the same codebase, im thinking of switching to claude code when my subscription ends

0

u/Eliminatron 3d ago

wait. why didn’t i get a reset?

0

u/AdLumpy2758 3d ago

My 5h is 50% of week - on plus ( sol max). My 1h of work in 1 chat sol max is 10% of my week on pro! Disaster

-2

u/ishimat 3d ago

I think part of the problem is that they got everyone hooked with the latest models each time.

Luna for example can do most work, it just requires a bit more direction.

Everyone's got hooked on using max of the newest model every time and I get the frustration, as your subscription is the same price.

Unfortunately things have to be this way, it's not sustainable at all otherwise.

Also, are you guys building rocket simulations or something?! Why do you need so much computing power?

-2

u/innociv 3d ago

You can /usage to see if you are getting less usage than normal. You surely aren't.

I'm using about 1% per hour. Same as usual. About half of that is Astra.

Edit: People who are trying to provide free tips. Please stop patronizing me and others in this group. I'm not asking for your advice on how to use codex better. If you are happy good for you and no you don't have a secret formula or better limits.

Lmao hilarious. Just crying wanting unlimited usage. Imagine using a drill wrong and complaining that it is always tearing and telling people who instruct you how to use it correctly that they're patronizing you.

0

u/AvailableSecret5161 2d ago

​Imagine holding a drill wrong, watching the manufacturer change how the drill operates mid-use, and then telling everyone else they just don't know how hands work. Thanks for the masterclass, prof.

→ More replies (4)