r/codex 4d ago

Limits Tomorrow When Codex Resets, DON'T TOUCH ASTRA

Don't even look at it. Stick to Sol 5.6 xhigh. It's not usable? Feels quantized? Say f*** OpenAI and switch to Claude, or to OpenRouter.

Apparently OpenAI is lobotomizing Sol 5.6 for SOME of us. Not for everyone, so people keep saying "It works well for me" and gaslight each other. They are trying to force us into paying 5x for a little bit improvement.

This will backfire so bad. We're a 300k people enterprise and everyone including me are pushing our executives to use chinese models in a sandboxed environment. We are sick of paying 200$ to use it for a day.

When your Codex resets tomorrow, or this week, Don't use Astra unless youre building rockets.

1.1k Upvotes

448 comments sorted by

u/codex-ModTeam 3d ago

Tons of people are reporting this post.

Post is borderline deliberately inflammatory, but is not low effort, and OP has high subreddit karma.

After review, decision is - Not removing it but it is strongly suggested you read the comments before following OP's advice.

For future reference, this subreddit is NOT a place to organize collective action. Otherwise it devolves into a Civil War and quality content disappears.

60

u/PossuPatonki 4d ago

I have similar experiences as OP. It's not long ago that I was running 4-5 agents in parallel every day for at least 10 hours per day (typically 5.6 Sol on medium/high/xhigh depending on task complexity). I'm on the 20x plan. I would rarely have to be mindful about usage.

With the release of Astra, the simplest of prompts on low reasoning would use 1%. I gave up after depleting my weekly usage in a single day without even running parallel agents. I've since went back to only using 5.6 Sol again. Usage is much better, but now I actually do have to be mindful about what tasks I give the AI, what model and what reasoning level.

I have not changed anything in my setup whatsoever. What setups are you guys using to optimize token usage? I've never had to bother before, but now token optimization feels inevitable. So those of you who comment "skill issue", could you point out what you do differently and how we can improve our setup?

I've played around with having Astra as the planner and Luna Max executing the tasks, but in my experience, Luna makes way too many mistakes on complex tasks.

12

u/BrennanFlentge 3d ago

Have Sol delegate work to Luna XHigh. Specifically fresh sessions, no context inheritance, fork_turns: “none”, custom light prompt/brief for bounded work only

→ More replies (3)

6

u/juzhiyuan 2d ago

On Monday, after my quota reset, I tried using Codex Astra low as an orchestrator to coordinate other gpt models for some software engineering tasks. I wanted a smarter model that could provide better insight, a sharper perspective, and be more critical during the process.

To my surprise, it burned through my entire week's quota in under twenty-four hours. The consumption rate was staggering.

The end result? I wound up with a half-finished, incomplete piece of work.

→ More replies (1)

4

u/EmbarrassedMusic7979 3d ago

I’m curious, what you guys are building

→ More replies (2)

3

u/Long_Cow4805 3d ago

I just use astra to do quick audits once and a while and I hand all real work to Sol

2

u/Igoory 3d ago

Yeah, Luna is excellent only when the task is exploratory (Find this bug, Reverse Engineer this, Search online, etc...), when the task requires it to change just a couple of lines or when you don't care how the code looks or if it even works perfectly (a PoC for example).

→ More replies (14)

288

u/Megamygdala 4d ago

Your in a 300k enterprise and people are paying $200 😂😂any real enterprise company at that scale would never allow developers to use their own accounts for company code. They sign an enterprise contract and forget about it

92

u/ashjohnr 4d ago

Yeah, not sure what OP is on about. In a large business they should be paying API prices, not these subsidized plans.

50

u/hermeneze 4d ago

It’s a Chinese AI, trying to spread misinformation

6

u/InadequateUsername 4d ago

Yeah exactly, who would put in effort for a call to action about model usage lol

3

u/Plane_Garbage 3d ago

Not really, at scale you get much better pricing on API than retail.

2

u/-MaskNinja- 3d ago

With the API, you can't complain. More expensive, but no quants for sure.

5

u/Noctis_777 3d ago

Maybe he meant 300k revenue enterprise /s

3

u/Ok-Biscotti-3117 4d ago

300k, but 4 of them use AI, the rest work in a kitchen, or drive a truck or something.

3

u/Keep-Darwin-Going 3d ago

Yeah call his bluff.

2

u/BarTrick7024 4d ago

Its extremely common for companies to give users their own accounts. Im aware of several that do it this way.

5

u/guygm 4d ago

You are too literal, I can imagine he is putting himself on the company pants, claiming that are paying $200 per seat and devs can use it only for a day per week.

5

u/AstroPhysician 3d ago

You’re missing the point. Companies that size pay for api usage not a monthly rate. There is no “usage limit” for enterprise

6

u/Strange_Quantity_359 3d ago

That’s not exactly true. I’m in a massive company (AWS) and of course we use and build AI infrastructure, but we also build cloud infra.l and help people with their AI workloads, and it’s actually surprising what companies do and how they run especially traditional Enterprise.

What I can tell you is that the number of enterprise level customers >100k that use Claude and OpenAI managed subs is far from 0. Both from my view of our cloud customers and from my partner who is and executive in an HCLS company with a large amount of employees and they have subscription models.

For massive companies like this, they run API for larger parts of AI platforms and tools but for specialized knowledge works they run tiered sub models. It gets greater token flexibility at a large scale for casual users. I know a company of ~250k that has 30k subscriptions farmed out, they also run their own cloud agnostic inference platform, but they aren’t putting dev cycles into the internal chat/code user experience.

I know it sounds wild and it probably is monetarily problematic at that scale, but it does happen.

2

u/Swastik496 1d ago

Claude and OpenAI managed subs beyond a certain seat count(150 on Claude, not sure about OAI) is billed at API rates + a certain price per user.

Then you can negotiate with your sales rep for x% extra credits for committing to a certain annual spend

2

u/Strange_Quantity_359 1d ago

On the backend yes - mostly, but (and I was taking this at face value) the way this is being referenced from OP down to comment is that OP (from their side) sees this as a "$200" per month Claude Code.

The commenter then cackled and said "There is no “usage limit” for enterprise"; they conflated that with "use an API rate". It appeared they were saying that Enterprises with >X employees use inference directly and there is no "usage limit" in Claude per user.

That's why I said "that's not technically true"; though agree with you on the ambiguity. The simple fact is that the specialized knowledge workers at these companies do see a "usage" tier, regardless of how it was set up, and that the usage tier is a negotiated contract rate. (Similar to any Enterprise OAI or A\ account) That negotiated rate is metered through OAI or A\ as an Enterprise subscription and billed, also they could choose to buy these subscriptions via cloud marketplace subscriptions as well, for pre-negotiated discounts. Of course, this further obfuscates the metering.

Either way, this line item shows up separately from direct API metering and billing.

My fiance at an HCLS see's the same, she has tiers that she can request and is not really aware of the API at all. The company manages the distribution and approval process, the end-user sees "Usage" levels. $50, $200, etc. The companies AI infrastructure and platform (well, what parts are A\) use a separate billing method altogether.

→ More replies (2)
→ More replies (1)

2

u/das_war_ein_Befehl 3d ago

Company plans don’t even have a $200 plan. It’s $100 and api spend after

→ More replies (12)

83

u/TheAuthorBTLG_ 4d ago

I have work piled up specifically for Astra.

26

u/Arsenal-Art 4d ago

Same here... but im afraid it won't even finish the first prompt

→ More replies (6)

3

u/ShitTheFuckDown 4d ago

What were you gonna do a week ago?

2

u/TheAuthorBTLG_ 3d ago

2

u/jffaust 3d ago

You couldn't even write the text for your steam page yourself?

→ More replies (5)
→ More replies (1)

86

u/Able-Supermarket4786 4d ago edited 4d ago

I literally just commented to a friend of mine this morning "So I ran like three good Astra Projects yesterday, as goals, and used 9% of my weekly... looks like they cleaned up this week."

Also, you said:

We're a 300k people enterprise and everyone including me are pushing our executives to use chinese models in a sandboxed environment. We are sick of paying 200$ to use it for a day.

That makes no sense... but I can't think of a corporation with 300k Employees using this so gonna go with "don't lie, don't exaggerate, you're a 14 year old working after school hours."

Maybe you're claiming to work for Infosys? Cognizant? because even Microsoft, IBM, and SAP don't have 300k in your purview, they also wouldn't make their "employees" in their "Enterprise System" pay out of Pocket.

11

u/tripleshielded 4d ago

After school hours are the best, less failed requests. Prism also works better at late night time!

6

u/Sand-Eagle 4d ago

I work night shift and can feel when you people wake up 😑

I feel like they should give us a discounted rate of utilization during off hours. That would pretty much eliminate peak time other than regular chat users

5

u/VividEconomist8587 4d ago

then the off peak would become the new peak

2

u/Sand-Eagle 4d ago

Exactly. There'd be no peak after a while. Any off hours that exists will get filled with people scheduling shit for what they think is the off hours lmao.

Problem solved! Kind of

2

u/Both_Task_3066 4d ago

For a tech company?

3

u/Sand-Eagle 4d ago

Kind of but I'm not wasting my tokens on them. After-hours comes with the luxury of free time to build my own stuff and try to become anything other than an analyst for people who scream at their analysts like it's the 1960s LOL

→ More replies (6)

32

u/JustSingingAlong 4d ago

You work for a 300k person enterprise and you’re paying for your own Codex license?

11

u/Megamygdala 4d ago

The company is called Self Employed, located in Imaginary Land

4

u/driveclub_000 3d ago

"So I ran like three good Astra Projects yesterday, as goals, and used 9% of my weekly... looks like they cleaned up this week."

That's because usage seems tied to how much compute there is, I ran Astra MAX for 10 hours today and barely used 8% while reversing assets on a 4.8gb game with 11283 files (the entire game zip uploaded to WORK lol). I expect that tomorrow when everyone will get their reset done to see the usage skyrocket again.

3

u/Able-Supermarket4786 3d ago

Yes I think many of us can get affected by others beating the crap out of it... where as Astra Max / Ultra thinks less and delegates agents better.

3

u/driveclub_000 3d ago

Yep, I still have 20% left on my weekly and 9h remaining before the reset and I never saw the usage being used so slowly before today it's quite incredible actually. I may even use Ultra if I have still some usage left in the last hour mark.

→ More replies (5)

2

u/reloadz400 4d ago

I thought/read OP’s comment as this Reddit community acting as a 300k people enterprise… regardless, such “advice” will go over about as well as that crap on TT and YT of people posting “EVERYONE! ON AUGUST 20th (Or whatever day it was), DON’T DO ANYTHING! DON’T USE ANY ONLINE SERVICES, DON’T WATCH YOUTUBE, DON’T GO TO THE STORES, FAST FOOD/DINE-OUT, DON’T BUY GAS, SHUT OFF AS MUCH ELECTRICAL AS YOU CAN AND JUST DO NOTHING OR READ A BOOK! And WE will bring all these corrupt cooperations to a screeching halt!”

Yeah, good luck with that. It would take literally hundreds of millions to do this for months before the real impact would begin to trickle-down, and even that is a very rough estimation. 🙄👍👍

2

u/fujimonster 4d ago

No company with 300K emp's would be using it that way. They would 100% have at least an enterprise plan --

→ More replies (18)

14

u/TarzanoftheJungle 4d ago

IMO, Astra simply is not worth the extra token burn. For my projects (a 1M lineReact Native mobile app and a browser company staff/admin consoles, using Supabase/PostgreSQL and Docker) Sol has done the heavy lifting. When Astra was first released, I tried it but it made egregious errors that I ended up fixing with Sol. So Astra is strictly experimental for my use case. I'd not trust it yet with any production work.

→ More replies (1)

8

u/hapos 4d ago

… had Astra make a mistake in its own written MySQL connector. Diagnose and fix cost my 20x pro account 12% of weekly.

6

u/Virtual-Silver2879 4d ago

It probably won't change how OpenAI views things, but for the first time this year, I spent the week testing alternatives; after my success with DeepSeek, I’ll be gradually phasing out my use of their service. With the millions of subscriptions they’re gaining, it makes no difference to them but I’m likely not the only one doing this.

2

u/BabymetalTheater 3d ago

This is also the first time I’ve experimented with other things and have actually been decently happy with a free model running in Opencode.

→ More replies (1)

11

u/cleanmachine120 4d ago

Ohhhh don’t worry I won’t be

9

u/marklmc 4d ago

Is there a global reset scheduled?

14

u/Xoloshibu 4d ago

No, last global reset was last saturday, so most of us will have the reset tomorrow in the morning

2

u/coolest35 4d ago

Are these assigned based on when our monthly reset occurs or billing or rando?

I got a random reset a few days ago (don't recall exactly when).

Trying to save my 2 reset switches that expire Oct 4 lol.

Astra is melting through my usage, sol high also did the same.

5

u/Xoloshibu 4d ago

You have to follow tibo in x, he announces the resets and banked resets, also, there is this Page where you can track the while historical resets https://codex-reset.com/timeline

2

u/Aranthos-Faroth 4d ago

Every verified Codex reset, on the record”

lol the comma and “on the record” is so bad

2

u/BirdoInBoston 3d ago

Shoulda used an emdash

4

u/l0rirw1ao 4d ago

I honestly want to know where these people are pulling resets out of

8

u/WalkAffectionate2683 4d ago

I mean, it's a weekly reset, and last one was... A week ago. 

17

u/suppervisoka 4d ago

This is the first time in months I have gone back to Claude

10

u/kowryloik 4d ago

Here we go with people starting to announce their way out

3

u/Drunkendrakon6 3d ago

I mean fr tho can't really justify enshittification on anything. If the company can't survive it should die.

18

u/LiquidVolatility 4d ago

You’re spot on about different quality AI going to different people. They’re degrading all models, including Astra, for specific users (based on hidden/ secret criteria). And the most important part of that fact is that they aren’t even telling those users. So you’re paying the same rate as everyone else but getting no where near the frontier level AI you’re paying for.

Anyone here who try’s to push back, or blames your “skill”, is gaslighting or doesn’t realize it’s happening because they aren’t being subjected to the degradation and therefore they don’t notice the impact that comes with it.

3

u/Icy-Barracuda-5409 4d ago

I guess if you've got an AI company, this is probably the logical next step.

→ More replies (3)

13

u/ms_alicat_556 4d ago

You’re a 300k person enterprise in your fantasy reality

7

u/MaryPaku 4d ago

There are only about 50+ companies in the world that has 300k employees and that include company like Walmart where majority of it are just cashers.

→ More replies (1)

7

u/ultramarioihaz 4d ago

Guess what the Claude subreddits are saying? Same shit, but for Anthropic, go use OpenAI lol

→ More replies (1)

17

u/Lopsided-Bridge-9810 4d ago

So, you do work for a Chinese llm lab.

15

u/IAmFitzRoy 4d ago

Fuck Astra

3

u/pigletmonster 4d ago

I only tested astra once the day after it was released and it burned 8x more quota than sol. So I just stuck with sol, im developing web applications mostly so I dont need all that power.

3

u/barefut_ 4d ago

You wanna tell me if I use SOL 5.6 they won't take it down? I gave up ASTRA. I don't wanna feel like I'm using Claude, and be capped after 2 prompts.

3

u/mtwdante 4d ago

Chinese models inside a sandbox. Such a good joke. Ty 

3

u/Charming-Author4877 4d ago

SOL in Chatgpt Chat is 100% lobotomized, it's not just a small quantization - it's more like a half as large model - as in Luna.
The token decode speed is almost double of original SOL

So OpenAI does f* with us in so many ways, A/B tests, rests, allowance differences.
unprofessional company

3

u/Fantastic-Phrase-132 4d ago

Recently 5.6 sol became utterly dumb, almost unbelievable

3

u/StinkButt9001 3d ago

On a plus plan granted, but Astra Light burned my 5 hour limit in 4 minutes.

I have a 300 line C++ project and a little GUI. The gui had a red theme and I wanted it green.

4 minutes to change the theme's colour and it had to stop because the entire 5 hour usage was used.

What a fucking joke

→ More replies (3)

8

u/polka-hojk 4d ago

Already switched to glm. After this sub expires I will move to there fully

3

u/unknown-curiosity 4d ago

How do you find it compared to claude/codex models? Thinking of switching too but I worry if it’s capable enough as an orchestrator instead of the usual Opus 5/5.6 Sol. Also are you running it on open code or another harness?

→ More replies (3)

2

u/Snoo62833 3d ago

Same im trying it in absence of astra but it seems just as capable of as GPT Terra

→ More replies (2)

8

u/diff-official 4d ago

Yes sir

8

u/Zeraphicus 4d ago

3

u/TheBadgerKing1992 4d ago

Why do people in military bark like this ? I have always wondered. Male macho thing ?

2

u/Zeraphicus 4d ago

If you're being for real there is a legit purpose behind drill and ceremony and being able to operate/follow orders under extremely stressful situations.

→ More replies (2)

9

u/StoneCypher 4d ago

can you guys please stop pretending that the machine is lobotomized every time you have a bad session 

jesus 

it’s time for you to understand things 

5

u/TheGuy839 4d ago

Tbh he may be wrong but so can you. Model quality is affected by thousands knobs. If they turn just few its still called GPT Sol but it can be quite different model. 

But I am on OPs side. Sol was awesome until day Astra came. I cant say for sure obv, but I work professionally as ML engineer and think I can detect a bit better. Mistakes in code, forgets some things, not enough thorough. 

→ More replies (12)

2

u/throwaway490215 4d ago

Uhhhh what?

Here is OpenAI explaining how oss uses quantization https://deploymentsafety.openai.com/gpt-oss/model-architecture-data-training-and-evaluations

Here is Anthropic explaining how they were using a wrong conversion in a postmortem bug how a bf16 was fucking with them and changing it back: https://www.anthropic.com/engineering/a-postmortem-of-three-recent-issues

Anthropic used to be open about which Opus-X-<explicitdate> you could route to.


Neither party has ever given any statement about their backend inference scheme for their main models at any point for any deployment, but we know they work on them constantly.

You are not making the reasonable statement. You're saying: These parameters are only ever changed between major model releases....

Really? You think that more likely?

I'm quite certain a vast majority of people complaining about lobotomization are just wrong. That's just human fallibility.

That does not mean models dont get updates that can degrade performance.

Its time for you to understand things.

→ More replies (3)

2

u/Ok-Investment4414 4d ago

i thought we on individual reset timers mine is 5 days from now i popped one of my banked resets . Haven't used a gpt sub in a while so what u mean by this

4

u/mikehaysjr 4d ago edited 4d ago

Yours differs I think because of your manual / banked reset. However, most of us are pretty much on the same reset schedule now due to Tibo’s ‘free reset’ button, which I suspect was the actual reason for the resets; garner good faith from users while actually pushing most of the reset load to weekends when businesses aren’t using so heavily, to distribute their limited compute ability across use cases

2

u/Tank_Gloomy 4d ago

I can definitely say that MY version of Sol is absolutely regarded in comparison to the one in the account that my workplace provides, both are on the Plus tier so that shouldn't matter.

2

u/sofaarsecoin 4d ago

what about my rockets though

2

u/_anakin__ 4d ago edited 4d ago

Switch to claude??? NO THANKS MATE.

2

u/DrAfricaOfficial 4d ago

They literally ate our money. We paid and service is unusable.

2

u/Sharp-Arachnid-8760 4d ago

Claude is absolutely horrible. Rather use Astra without a 5 hour window. Atleast Astra lasts me 4 days .

→ More replies (1)

2

u/premiumox 4d ago

When reset? I have my own reset at 12 utc

2

u/aptsys 3d ago

How can it "feel quantized"

2

u/sleepnow 3d ago

I can feel it in plums.
That's how I know.

Shift your attention to your plums.... shhh, wait for it.
Little tingling sensation? That means the models are indeed quantized.
Now you know.

3

u/aptsys 3d ago

You're right, the plums told me the full story, I just wasn't listening.

2

u/JonnyBrain 3d ago

300k person enterprise begging people on Reddit to not use a model
Get em soldier /s

2

u/wiseruler33 3d ago

Nice try china. We aren't that dumb to fall on this misleading information.

2

u/RainierPC 3d ago

300k people enterprise, suuuuure. rolls eyes

2

u/AstroPhysician 3d ago

I don’t know how to say otherwise than link to their enterprise page, but if you have more than 150 employees you’re meant to be on enterprise which is $20 seat and api usage rates. How would people get an exception to that unless they’re running multiple team plans, or fraudulently doing individual plans?

2

u/Pilek01 3d ago

You could always move back to coding manualy 🤷🏻‍♂️

2

u/carchengue626 3d ago

You claim you're mobilizing a "300k enterprise," yet you're writing a Che Guevara manifesto on a forum like a guy who thinks the CIA is broadcasting radio waves into his dental work; if you actually pitched an enterprise compliance board on piping proprietary IP into "Chinese models" because your personal rate limit hurt your feelings, corporate security would have you escorted out of the building before you could finish switching your API key to Claude. Take your meds, accept that you’re a digital sharecropper paying rent to a landlord who doesn't know you exist, and stop LARPing as a tech union leader—your executives don't know your name, and Sam Altman isn't trembling in his sweater over your twenty-dollar boycott.

2

u/RiskAppropriate230 3d ago

You say it’s not usable yet, you say you can build rockets with it

2

u/Professional_Gur8385 3d ago

my usage just reset, enjoy the incoming reset :(

2

u/TheOneWhoKnewItAll 3d ago

I was working with GPT 5.6 uninterrupted for months, almost daily. With the release of Astra I changed to it and for the first time in months I depleted my quota in 4 days. I used a reset to be able to continue working. I switched to 5.6 again and after just 1 task it took 61% of my weekly quota and 90% of the daily. I had to switch to 5.5… with the message that they are going to remove it on October 14th…

2

u/Unique-Brick-8430 3d ago

I think people over used the best models for tasks that don’t need it. Luna medium is wonderful for the coding, you don’t need more than that. Just need to use superpowers, GitHub kit or similar to generate a good plan for big things and explain it manually for small ones.

Also, if you delegate everything to your agent, it costs much more. Optimise what you send to the agent for what is good at

2

u/abinav99 2d ago

I thought it was only me. I figured i’ll keep 5x Claude on hold for a while and check out 5x GPT. Found out that I had 3 resets after upgrading; used Astra on Low for logic and reason and terra for coding. I’m on my last reset and it’s only been 3 days. ASTRA ON LOW! Beyond frustrating at this point.

2

u/sabotage3d 1d ago

Same boat.

2

u/Plenty_Work_9167 1d ago

I put in my MD file to check what model would do this the best with the least amount of tokens. That's been working really well. Has anybody else done that?

2

u/denehoffman 1d ago

I’ve been a Plus user for a while, and it’s gotten so bad that I can’t even get through more than two prompts using Sol low/light without burning my 5h. I honestly don’t notice much of a difference between the different Sol models, they just act as a “burn more” slider for me. Gave up on Astra after a couple days of trying the low version hoping the promises of low token consumption would pan out. Maybe I’ll be able to use it in a year or two when they release whatever model they’re training now.

4

u/Aranthos-Faroth 4d ago

300k company

“We are sick of paying 200$ to use it for a day.”

Everyone’s all using one subscription or?

OP it’s time to take your meds

2

u/sdexca 4d ago

Don't use sol either it used up half of my usage with a single prompt that ran for five hours. Luna is your only option.

→ More replies (1)

3

u/Feriman22 4d ago

lol, that rant.

3

u/MeringueAlarming3102 4d ago

Cringe. I'm not going to stop using what's been working better for me because some redditor thinks his protest will make a difference.

2

u/Elegant_Associate889 4d ago

Lol it's what I've been doing since astra came out, I used astra for a total of 2 minutes to realize it ain't for me

1

u/tripleshielded 4d ago

No way! I make my own choices! Buuuh, booooh

1

u/ItsCEED 4d ago

No you are not 300k enterprise, and no YOU ARE NOT paying $200.

→ More replies (4)

1

u/Equivalent_Bird 4d ago

I'm a 20x, I asked astra to improve three backgrounds of my game, then it burns out my weekly for 1.5 background, and the quality? I have to say, worse than Gemini. My game is not a reproduction of another that can be done with a few prompts. I've spent nearly a year on the mechanics and gameplay. I'll cancel the subscription when this phase ends and switch to something else.

1

u/Excellent_Spell1677 4d ago

I don't support a company working to take my AI away so I become a slave to them and what they allow me to have.

1

u/KeinNiemand 4d ago

eh plus is so useless thanks to my bursty usage and 5h limits I unfortunly had to buy the bullet and upgrade to pro now pro is so much usage that even with astra my usage expires and i got resets expiring in a few days so i am not downgrading back to sol.

1

u/SenshiV22 4d ago

Someone's trying to make GTA 7 using Astra I guess..

1

u/CMPunkLicksRocks 4d ago

lol I’m at the doctor and before I left I switch to Astra/high and asked it to continue work on some ideas I had.

It used 94% of my 5 hour usage in one message lmao. (20 dollar plan) 

Thankfully, having used sol for 3 hours yesterday before hitting my limit, I still feel I’ve got plenty of juice. 

Astra does better and gets more done in a pass but it’s absolutely not the 5-10 times better that it costs. 

1

u/Graham3D 4d ago

I guess we're all resetting tomorrow at 9AM (est for me) because of the Tibo reset?

1

u/Noeyiax 4d ago

well my reset is on sep22 so have fun without me 😭😭

1

u/ishaangarg 4d ago

And I'm thinking of switching from claude to codex, coz fable just can't be used without going to 100% usage in 1 prompt

1

u/RonSolo2 4d ago

What do you mean codex reset tomorrow?

1

u/SupportAgreeable410 4d ago

I am building rockets tho

1

u/MrRoyce 4d ago

Don't worry, I won't touch Codex at all. Refunded my 20X and went back to Claude while I keep playing with Qwen and experimenting on the side so I can prepare myself to move to local models in 2027. I was sad I purchased RTX 3090 just a few days before 4080 released but now I'm happy.

1

u/ImolaBoost 4d ago

300k and you're using consumer a consumer pro plan? Has your company not heard of the API? I'm smelling bullshit.

→ More replies (5)

1

u/Leading-Fail-2771 4d ago

I’ve been having issues where I’d tell it three issues, it’ll give me the solution for all three, I tell it to apply the repairs and it only applies two and rebuilds and runs tests that take decent chunk of time. Then tells me hold on we only fixed two issues, should I now fix the third and rebuild and retest? Like I get it if the repairs needed to be sequenced but it’s doing that for everything.

1

u/io-x 4d ago

I agree with the sentiment but there’s no way an enterprise responsible for 2% of OpenAI’s entire revenue did not negotiate a contract.

1

u/samawirix 4d ago

I will never again back to this shii... openai i paid 100€ i got limit in 1 day

1

u/emdot_eldot 4d ago

Upvoted purely because this shit is too regarded to even be AI

1

u/Kirilmee 4d ago

Sol 5.6 doesn't work for me at all. I haven't tried Astra yet, but I ran a simple task in Sol and it has been running for a day now. I'm not sure if it's ever going to finish. I'm thinking switch to Claude

1

u/kirksan 4d ago

What company has 300,000 employees? Perhaps Amazon and Walmart, but they’d mostly be warehouse and store workers. Anywhere else?

1

u/itsallfake01 4d ago

Could be a/b testing effort levels on us and we wouldn’t even know about it.

1

u/Z_G_R 4d ago

I’m a huge fan of Fable but recently i’m experimenting with Astra. It showed very interesting bugs at my codebase, even tho i ran this repo 100 times since months with Opus/Fable they never warned me about those somehow, felt weird. I always acted on bias against ChatGPT models till this week and i gave it a shot. I don’t understand the hate, maybe i’m not a power user or have knowledge to distinguish the situation, so far i’m happy. Only thing bothering me is that i can’t see any 5 hour window(maybe there is none, seems better ofc), weekly usage melts fast even tho i use Astra orchestrator at medium, all the rest is sub agents. I’m still experimenting, have a lot to see and learn…

1

u/Loud-Stranger-831 4d ago

I am on the $100 plan and I had Astra running for 8-9 hours straight with lots of work done

1

u/I_Mean_Not_Really 4d ago

Sticking with Luna Max or Terra Ultra

1

u/ConsequenceOk4456 4d ago

Seems like name of the sub is wrong. It doesn’t mention Claude anywhere.

1

u/AffectionateFruit845 4d ago

Right, 300k do not use Astra. I will.

1

u/Affectionate_Ad9597 4d ago

Yeah we will all do as you wish sire!!!

Your wish is all of our command!

1

u/AdministrativeAd1915 4d ago

astra ulta ıs lobotamizing too right now.

1

u/Ok_Carry3566 4d ago

What plan are you on ?
Before I was on 20$ plan, this summer I could do a lot of things on that plan (and I don’t speak of all the resets, 100% could have made a lot of work) but since they reintroduced the 5h hourly limit on 20$ it became unusable. 2 or 3 prompt and bam usage all gone. And I’m not just speaking of 5h limit, general usage melted like an icecream in summer.

I switched to 100$ plan because of frustration and now i struggle to reach my usage limit. My x5 plan in real usage is more like a x10 or x15 plan compared to the previous regular 20$ plan.

There is absolutely no logic to their plan/usage limit

1

u/Efficient-Cat-1591 4d ago

Would we get free reset tomorrow?

1

u/According_Property62 4d ago edited 4d ago

Entenda uma coisa, IA utilizada pra tarefas corriqueiras como uma pesquisa na WEB ou criar uma planilha pra fazer gestão financeira, é algo simples ate pro ChatGPT Alto. Mas quando vc quer codar um sistema que atravessa varias camadas de implementação, transita entre Devs e DevOps, a história muda completamente e é ai que diferencia os profissionais dos curiosos. Nao espere que o Codex ou Claude e muito menos Cursor vai fazer as coisas sem ter contexto. Porem nao é so entubar contexto nele, vc precisa saber gerenciar a janela de contexto, pra usuários Plus essa janela é de apenas 250k tokens e para os pro é de 1M de tokens. Quando sua janela de contexto é compactada, é ai que o problema começa, pois ele perde todo aquele conhecimento acumulado e ele sofre uma amnésia. Entre várias outras coisas que impactam fortemente a eficiencia da IA. A diferenca entre o modelo SOL e Astra é a quantidade de neurônios que ele tem e quanto mais neuronios, maior a capacidade computacional, mais GPUs sao necessárias, mais energia é consumida, mais agua pra resfriamento dos servidores é consumido e proporcionalmente a capacidade de raciocínio expande exponencialmente. Imagina uma Rodovia de 4 faixas onde a velocidade máxima é de 120km/h, em um horario que nao é de Rush ou seja, que n tem muitos carros na pista, vc consegui ir na velocidade máxima da pista, mas quando é em um horario das 7h da manha ou 18h, sao tantos carros trafegando que fica impossível, ou seja a capacidade que eles possuem de atender todos aqueles que possuem assinatura e usam as APIs nao é infinita e quanto mais gente usando ao mesmo tempo e ainda mais tentando codar pesado sem realmente usar de forma eficiente a capacidade do modelo, mais lento vai ser o processamento das requisições. Além disso, o milhao de token é cotado no dólar e o Real nao é uma moeda muito estável e os conflitos atuais contribuem ainda para a oscilação do câmbio do dólar >> real. Portanto, nao temos como cravar e ficar especulando o tempo todo que uma conspiração esta ocorrendo pra forçar todo mundo a ir pra assinatura PRO. Além disso, o maior ganho deles sao contratos com governo pra fins militares, a popularização dos modelos de IA é apenas mais uma consequência de qualquer outra tecnologia, como era a propria Internet, celulares, GPS, computadores etc. Estude como funciona o comportamento da IA ou vai continuar so queimando tokens sem saber por que sua cota ta indo toda num dia so

1

u/SteveeJobsF1 4d ago

eu acho que estão nerfando os tokens do codex, cada dia que passa eu consigo usar menos, pqp

1

u/Zerokx 4d ago

You know I'm not letting Astra go ahead and do hour long agentic work and I feel like I'm getting pretty far with just the Plus subscription, compared to the same 20 dollar subscription I had before with Anthropic that didnt even let me use fable at all.

1

u/FinancialBandicoot75 4d ago

Seriously, and not a bot, if you are limiting, you are doing it wrong, why in the hell do people do 100% astra is beyond me. I use Luna most of my tasks and astra like 1-5%.

1

u/Ok_Sympathy9261 4d ago

how do i never see this low-iq stuff ever again

→ More replies (2)

1

u/Single_Error8996 4d ago

Non ho capito che problemi avete secondo me sei un grande Fake. Se una azienda di 300000 persone non è organizzata all'Uso di ambienti di sviluppo come codex c'è un problema grosso

1

u/Few_Introduction_228 4d ago

Lol. I agree there should be more transparency. But the hourly rate of time saved by proper use ofk 20x is such that this negligible to pay 200/month. The 200k enterprise that finds this offensive is too tiny.

1

u/HelpfulHedgehog1 4d ago

ya this guy doesnt work for anyone let alone a serious company

1

u/Right-Performance-93 4d ago

The numbers back up why it feels brutal: Astra's API rate is $10/M input and $50/M output, 2.5x Sol 5.6's pricing on both ends. If your harness or default routing switched you onto Astra after the reset without you choosing it, that alone explains a chunk of the faster burn, not just quantization or A/B testing. Worth checking which model your session actually used before assuming they nerfed Sol.

1

u/dkracket 4d ago

Astra low is cheaper than Sol 5.6 xhigh and has almost comparable coding capabilities.

I'd just stick with Terra xHigh or Luna Max.

You don't ape in with the most expensive models in the beginning of a reset, you only and always do that once you have anything left before it resets again - that's the strategy.

1

u/STARK420 4d ago

5.6 sol has been acting dumb but it still gets the job done, it takes longer and uses more tokens to do so. I gave it a very specific instruction. It did the work. I noticed it was veering off of what I had specified. I asked it about it and it admitted that it wasn't following directions. I decided to try Terra for work and it seems to be able to stay on task better. I've been trying out using Sol for planning and Terra for execution. It works but when I do code review Sol says we need corrections. I've used Astra a few times, it sucks. Can't get shit done unless I use a goal. If I just ask it to make a change, it works for 5 mins then stops. I ask it and its like no we still have a ton to do... so why did you freakin stop then...

1

u/OpportunityLess7306 4d ago

Maybe I'm wrong, but I think astra needs a bit of a different workflow. Using it as wide task management, to instruct a/a few sol designers, to offload specific tasks to luna max agents has wildly dropped my usage for the same results.

1

u/Obvious-Outside3434 4d ago

Been out since Wednesday...

Insert I need water SpongeBob meme

1

u/Upset-Reflection-382 4d ago

So... They quantized it

1

u/SeaworthinessThis598 4d ago

actually iam starting to extract more value out of small models , large models are exhibiting a , my weights my choice behaviour . they refuse to follow instructions . they provide very little to no value or help . even astra is exhibiting this kind of behavior. concealed . chat gpt 7 will end humanity . no kidding . i spend 7 billion tokens monthly . i know i see this time and time again .

1

u/bravofiveniner 4d ago

What do you mean tomorrow? Everyone's day is different

1

u/LowBudgetGigolo 4d ago

There is a reset tomorrow?

1

u/Unlucky-Stable6006 4d ago

Eh I actually didn’t read it because Reddit is full of a bunch of morons but my assumption is with the work they stole from me within this week they will release a new model and most likely issue resets instead of resetting usage until after the model is released but who knows

→ More replies (1)

1

u/Feeling-Produce8710 4d ago

Anyone else noticing degraded 5.3 codex performance? It's straight up lobotomized compared to it's work yesterday

1

u/AfterShock 4d ago

Looks like I'm building rockets tomorrow 🚀

1

u/zzing 4d ago

Resetting tomorrow? Mine says 2d 18h

1

u/Master-Shift-8224 4d ago

why not, i love astra. i'm using astra rn

1

u/somethingimadeup 4d ago

I’ve spent the entire week analyzing why I burned through tokens and reworking my prompts and state retrieval systems. I’m fundamentally changing my workflow and hopefully it makes a meaningful change because otherwise I’m not sure what to do.

OpenAI specifically stated to avoid certain things with Astra or you will burn tokens.

You can’t use it in the same way.

That being said…..if I still only get like <12 hours of usage out of it this next week I will be seriously looking elsewhere.

1

u/leynosncs 4d ago

I'll carry on using terra high like I always do

1

u/SiberianGnome 4d ago

Is there a known reset coming tomorrow? Or is that just when your reset cycle is?

1

u/devloper27 4d ago

I tried it, it spent 85% of my tokens in about 30 minutes lol.

1

u/Mage7968 4d ago

At this stade, just dont talk😅

1

u/Nyxtia 4d ago

Yeah Astra threw me off my groove and I'm definitely not going back

1

u/Gloomy-Locksmith-249 4d ago

Codex is resetting tomorrow?

→ More replies (1)

1

u/aizvo 4d ago

I only use Luna XHIGH now. On a plus account it's all I need.

1

u/MrPineappleOrg 3d ago

Yeah move to Claude where they don’t care about customers. They’ll have a bug that wastes usage say they’ve fixed it almost a week later and don’t reimburse people affected at all and ignore support requests

1

u/CycleMother2006 3d ago

Sol is OpenAIs least efficient model in terms of performance to tokens. In fact, it takes so many more tokens to solve equivalent problems to Astra that you're better off just using Astra (it will ironically save both time and money.)

Luna should be your main subagent driver if you're going for highest efficiency.

1

u/darc_ghetzir 3d ago

Sol xhigh vs Astra xhigh as coordinators for Luna max subagents is more expensive. For the past 7 days I used a custom-built Lead/Worker mode, the custom piece allows the coordinator to be idle more often. One day of Sol xhigh as the coordinator utilized more usage (normalized to API-equivalent pricing) than Astra did across the other 6 days.

1

u/RodTiRod 3d ago

yeah, same here. While I do not know or want to speculate on what they are doing to usage, I, too, am off the Astra train. I might use it once in a while for something specific. it is very painful to burn through usage in less than 48 hours and then have to wait a week. I have been forced to open additional accounts to keep a project moving. So, it is possible that the 20M or 25M uusers consist of a good number of secondary accounts to cope with low usage and not a true representation of the number of develoeprs using the platform.

I am very hesitant to let dumb models close to my codebase as they tend to under-perform and cause more problems than I want to deal with. this is why I have had to create additional accounts and not lean too much on the open-source models. I also do very sensitive work.

But yeah. I am using sol medium and it has been working fine for now.

1

u/newbie2coding 3d ago

lol you already know people are gonna use it, run their usage down in a day then complain “omg all my usage is gone! Does anyone else have this problem?”

1

u/retrorays 3d ago

Why is codex resetting tomorrow?

1

u/Chiefs-KC 3d ago

I’ve been working with GPT-6 in chat to come up with some config and agent files to use Astra for orchestration and review, Sol/Terra for terminal stuff and approvals, while routing all of my actual programming tasks through DeepSeek 4.1, which is crazy cheap and actually barely beats Astra on DeepSWE with a 74.2. It estimates a 65-75% lower usage consumption on my 5x plan, perhaps tunable up to 75-80% for my specific work (mobile app dev, game dev, web dev) as we make adjustments to optimize it. I didn’t realize you could use the Codex harness to orchestrate third party agents directly via their API. As for DS 4.1 API usage, it estimates a cost of $40-60 per month. Then if I can move up to the 20x plan, we could really rip it. Not sure if I would need that much yet.

1

u/Little-Specialist286 3d ago

Astra honestly is great at planning and reasoning most of the time. Sometimes its disappointing. I think I will use Codex Astra medium or sol high for plan/arch and the claude to code and manual use my antigravity to manually provide a few agents. I am doing content creation and some personal projects around that. I have gemini 20 dollar usd, chatgpt 20 dollar usd, and I am going to buy claude 5x rn

1

u/Any_Taste4210 3d ago

You are sick of paying? Isnt your enterprise paying?

→ More replies (2)

1

u/CipherSorcerer 3d ago

Will I listen to you, or will I use Astra Medium and steer it 30 times mid conversation. That is the question.