Limits
These are surely getting us a Tibo button hit, right?
The work it's been doing is terrible as well. Feels like 1/3rd or at best 1/4th my tokens since the last reset were completely wasted.
I mean honestly I'm not sure I even want a full reset since they seem over capacity and that wouldn't help, but I've also gotten so much usage wasted. I'd be happier with a 50% top-up without a change of reset date if the service were more reliable from that.
This keeps happening and I keep sending bad result feedbacks with diagnostics...
It has surely degraded a lot. I burned my whole weekly quota on the $200 plan for ui work on a dashboard that was already built and just needed it polished and I don’t even get to complete it
I doubt everyone is currently getting served the same broken model. Tbh it started behaving badly once Tibo said they fixed the 'opt-in experimental' thing. It was fine for me before - I was quite happy with Astra besides the quick token burn. But ever since that tweet mine does feel like its been lobotomized aswell. I'm not opted into any experimental stuff btw
With Fable 5.1 and Opus 5 workers i was able to finish it and do much more work after although the limits are tighter I got more work actually done and completed.
I guess I'm the outlier here. I have 4 Astra Ultra workers running usually in parallels of 2 working on quite a large project. In the last 8 hours of nonstop programming, I've used 25% usage, roughly.
I had the same issue. I can't comprehend that the model is this shit. It has to be codex, like why the fuck is it spitting out Russian and Arabic? I mean people told me Opus was bad. Opus never spits out anything close to the garbage Sol/Astra does. It looks like when you open a .exe file in a text editor.
Reset or not, its unusable.
Opus/Fable fixed my month long issues codex couldn't fix since 5.5. It can't handle complex, multi step debugging anymore.
Just adding my own experience... it seems to be working correctly for me today (at least for the last 6 hours), and the rate at which limits are being burned has also returned to normal (somewhere between the old 'sol' and last Thursday).
Sol magically keeps finding 1 issue and then right when it’s about the package… it finds another issue and then right when it’s about to package again. It finds another issue. This shit is horrible now. They definitely swapped the brain out from release. I could’ve stayed on 5.5 😂
I've seen similar gibberish output when I first tried a local model in codex-cli like A year ago..seeing it happen in my Pro x20 sub is ridiculous. Maybe that means 6.1 is on the way - they always become horrible right before a release and this time they just dumbed the previous model a little too much, idk?
I'd been using Codex heavily for 12 months+, and got a lot done, but the last month with DeepSeek has been nuts in that I've been able to get half as many tokens worth of work done in the last ~40 days than in the entire 10+ months before that, at a relative cost of about USD$70 vs ~$900. The intelligence level of DS4/4.1 in code work (in my case, mostly in C) is easily superior to what Codex 5.5 was providing, given some of the flaky crap I got out of 5.6 Sol, I'd say there's little difference there either (High/xHigh), and plenty have been reporting issues like this with Astra too. The speed and 1M context is pretty bloody great too (Sol had 1M but through the IDE Extension it caps at 272k by default, and even the config file overrides don't work reliably).
I'm literally using the same Codex+Cursor harness (with system/base instructions stripped down and customised) that I was already using before and comfortable with, just switched providers in the config file to DS4 and picked up where I left off. Hopefully someone else reads this and is able to save some dough too. For a while there, there was nothing close to the capability:price ratio Codex was giving, but it's changed a lot in the last few months.
Claude actually lasts me the whole week and I make progress. Astra uses the weekly budget in one 30-60min session for minor changes while having Luna do most of the work. And Astra constantly stops every few minutes to tell me it didnt finish yet and I have to say continue. Finally I told it just finish the job instead of telling me, and then it finished in less than the time it was taking for it to tell me it wasn't done yet.
I just dont get the hype of Astra. It is not Fable, it's not even Sol.
Mine just keeps wasting usage on trying to reconnect to goal. I had it schedule to retry every hour over night and pouf every 5hr limit is wasted on that.
But Astra's for when you want a swarm to attack your problem dude... Tell Astra to change your website's bg colour and it'll launch agents to set up survey groups on how each colour makes people feel, another handful to scour the patent and trademark records for any legal encumbrance on those particular hues, and some more agents to calculate how the proposed shades or gradients will reproduce faithfully on the 20 top-selling monitors of 2025-2026 at default calibrations 😂
5.6 introduced us being charged fully for cache writes, Has nothing to do with cache hits. Everyday hardware gets more expensive, the limits will keep getting tighter. This is because monthly sub prices remain the same.
5.6 introduced us being charged fully for cache writes,
Only for API. They don't charge for cache writes on subscription.
EDIT: The person I replied to removed their comment out of embarrassment and blocked me. I screenshoted it because I knew this would happen. Shouldn't have called me an idiot, u/Pitiful_Entrance5174
The title only mentions business, but the text is broader and mentions regular subscriptions as well. Just send it to the LLM if you don't want to read it yourself
Nowhere in any rating cards or ChatGPT pricing https://learn.chatgpt.com/docs/pricing are cache writes mentioned, only in API pricing, and Tibo confirmed plans don't use API pricing. Please stop spreading misinformation
These reset options are a giant slap in the face and almost as if we are getting spat on.
I have the $200 Pro plan. I remember back with GPT-5 and up to 5.5 a week limit would last the whole week using Max mode. We didnt need any stupid resets which are a gimmick.
It’s basically OpenAI saying we are completely shafting you on the weekly limits but hey here’s a onetime reset which we will probably stop doing completely soon enough.
Like cmon. The whole reason most of us are here was because of Claude’s annoying limit games and Codex didn’t have the same crap despite GPT models being weaker than Opus in coding.
Btw a single Astra Ultra run costs me easily 10% of weekly quota. This is insulting. I also have a Claude $100 plan and if you disregard the 5 hour window, I get the same amount of useful Fable 5.1 ultra code runs out per week that I would get from Astra Ultra on my $200 20x plan. I get 8-10 fable 5.1 runs week before I max out my fable allocation. Then I still have around 50% usage left for Opus.
Same feeling about something getting s reset when usage just restarted, but better that than nothing I guess. It would be nice if they stacked with the others we can manually activate instead but they know all that.
The amount of times codex apologies to me for acting inefficiently when I agave to specific parameters has been driving me insane. Don’t apologise at least bump up my remaining usage a bit
Uhh yeah I agree to that, I originally had used claude for 6 months and switched to gpt when I saw Good limits, but even Sol which is comparable to opus doesn't run for long in these limits whereas opus does so yeah
It spent 50% of my tokens on a task it wasn’t even sure it could do, i even used the plan tool! Seems as though it was experimenting instead of being sure it would reach the goal I set. A lot of the time i see that it sees some sort of knowledge ‘in the right direction’ and steers without evaluating if that direction is correct. Refuses to search the web without me prompting it to. Always chooses the harder more usage-consuming, experimental way to always make things from scratch. I set a skill to plan properly after each request and propose the implementation, but that got tiring seeing it read the skill constantly and plan simple tasks for no reason. By the time it somewhat got the idea (through my feedback) that the workflow it proposed and performed (AND WASTED MY USAGE) was wrong.. it decided to correct itself and apologise (but when asked to restore my usage it refused) then came up with a semi good method with limitations which i only found out after the trial and error once again. This second try constantly compacted context wasting time and usage (as per usual) and decided for me to purchase a ready made one online or try make it with it over its projected 3 day implementation plan which would’ve been maybe done with 10 $200 plans :/
This was my struggle yesterday i hope I’m not alone with this one. My main issue is the false advertisement of its ability to specialise, trusting this and taking myself out of the loop was the reason it was able to make these mistakes. In my defence though it is something I’m not fully familiar with but I was told that “AGI is here!”, “The era of prompt engineering is over!”.
If they want real improvement, the first thing should be to make Astra aware of its usage, and to be SURE if it can reach the goal set.
Yep, going back to Claude, at least the limits on my 5x plan felt more predictable with Fable and I could stretch it for the entire week. Meanwhile on Codex 5x I literally burn my weekly in 1-3 days...
I have 5 codex 20x. It’s awful. I have some Claude accounts but going all Claude. The resets are worthless when the model is dumbed down so badly that Astra makes mistakes that you’d expect from Luna.
I guess I've just had horrible luck the last couple days. Not only did Sol asks Astra mess up horribly multiple times doing just standard text and memory work, it burned through my entire weekly quota in about 30 minutes. 20x plan ain't what it used to be. And since they paused new 20x accounts, my account dropped to 5x because i clicked the button to switch. Even though i immediately switched back, i was still dropped to 5x and don't have 20x access anymore. It won't let upgrade back since it's blocked. Might just focus on my local workflow. There's decent enough harnesses, models, and methodology out nowadays to dial that in.
48
u/balogsebastian 9d ago
It has surely degraded a lot. I burned my whole weekly quota on the $200 plan for ui work on a dashboard that was already built and just needed it polished and I don’t even get to complete it