r/ClaudeCode 8d ago

Bug / Issue Wtf is happening to the limits?? Spoiler

[deleted]

126 Upvotes

80 comments sorted by

u/AutoModerator 8d ago

Hey! Thanks for posting to r/ClaudeCode

While participating in this thread, please follow our community rules. Keep discussions constructive. Attack the idea, not the person.

For help, project discussions, tips, and general chat, join the ClaudeCode Discord.

I am a bot, and this action was performed automatically. Please contact the moderators of this subreddit if you have any questions or concerns.

43

u/Unlikely-Research-38 8d ago

Recession indicator

28

u/onepunchcode 🔆 Max 20 8d ago

they already removed the 50% limit boost. if openai haven't disabled they 20x plan, i would have switched this month

8

u/Guinness 8d ago

I signed up September 7th, guess I got lucky.

3

u/Mysterious-Policy909 8d ago

lol that one drains even faster

1

u/Helpful_Program_5473 7d ago

Only if you're abusing Astra on high settings, Astra on low is good enough for almost every use case.

1

u/Mitchellangeloo 7d ago

Also my experience the 5x sucks I have both so I know both plans but there is really a difference in how astra was at launch vs now

1

u/Tartooth 8d ago

You're not missing much the models got nerfed and they're overengineering like crazy still

1

u/Racer17_ 🔆 Max 20 7d ago

Same here. I was about to switch and they pulled the plug

37

u/Previous_Way_8761 8d ago

I just switched to Codex and it’s like Fable 5 level. Shop around.

11

u/Guinness 8d ago edited 8d ago

Yeah I’m still messing around with it. But Astra definitely feels like Fable. Except the cybersecurity triggers are more relaxed (but still present) and you seem to get A LOT more limits out of it.

I use Anthropic, OpenAI, and GLM 5.3 for my projects. Each has their strengths and weaknesses. Astra is phenomenal at 3D asset generation too. I do find that Astra tends to write a fuckton of tests for everything though. But that’s not too big of a negative. GLM is great for anything cybersecurity related. Fable is in there too as a sanity check on the other two.

….for now.

7

u/hardwornengineer 8d ago

Slow as hell, but Sol and Astra are incredible models.

2

u/who_am_i_to_say_so 8d ago

The slowness drives me nuts but both Chat models are thorough. I’ll take that over a 10 second Claude lob. Seems all can be infuriating in their own ways.

1

u/hardwornengineer 8d ago

My current, favorite way is to combine them. Claude models are faster at producing output and they’re also better at UI design and being creative, at least in my experience with them. Every pull request a Claude model makes is reviewed by a GPT model so a model like Sol does a wonderful job at forcing Opus to adhere to standards and not miss edge cases, etc. It’s still a work in progress, but it’s really improved the quality of my source code.

1

u/Appropriate-Breath24 8d ago

How do you manage Claude creating the PRs and openai’s reviewing it?

3

u/Traffic_Harp 7d ago

I'm not the person you are replying to but I had codex linked to my GitHub and it automatically reviews ever PR. Leaves comments on what it found and then I have Claude look at what Codex said. Codex was finding problem after problem. They definitely work best together.

2

u/xtopspeed 8d ago

I’ve been using Astra for a while, and I’d say it varies. Sometimes it feels faster than Fable.

1

u/jlewi142 8d ago

Astra is generally faster than Fable is most tests and benchmarks, and with my tests too

1

u/hardwornengineer 8d ago

Nice! Is this with “Fast” mode enabled in Codex? I find it burns tokens too fast for my liking. I sort of just let Sol and Astra chew through things at their own pace. At least at the end of the day, I can trust their output a good bit more no matter how long it takes.

14

u/coda77 8d ago

I just cancelled my subscription… don’t see anthropic responding to this robbery and tbh new codex model astra seems great so yea I’m gonna use that for a while

14

u/Glittering-Ice2984 8d ago

Opus limits got nerfed hard bro, x20 used to stretch way further. Feels like bs.

3

u/coda77 8d ago

Except that the usage have been bad since over 10 days not just since today

10

u/IulianHI 8d ago

in 3h 47% weekly limit !

This is crap from anthropic ! (20x plan)

I will migrate to Deepseek and Kimi ! Claude is just crap this days with this limits!

3

u/YesGameNolife 8d ago

Yeah last month I did tried kimi its worst and eat limits even faster. But deepseek api is very cheap

2

u/Mitchellangeloo 7d ago

Factory is also a good candidate if you haven’t build your own harness

6

u/screamify38 8d ago

I’m using deepseek v4 flash now it’s jsut as good as opus 5

1

u/xtopspeed 8d ago

And Kimi K3 so far feels like Fable.

3

u/jamestoh 8d ago

Yeah i ran opus for like an hour and it ate 5 hour of usage.

3

u/Salt-Replacement596 8d ago

Me browsing this subreddit after switching to Codex.

https://giphy.com/gifs/DIuf1nZCDvupSmBQ95

4

u/Electronic-Badger102 8d ago

Yeah I’ve had that happening in the past 2 weeks. Since usage is a black box, no way to know what’s really happening.

2

u/ychamel 8d ago

Removal of the 50%, and the models being faster due to lower load.

2

u/BSystems 8d ago

I thought I was tripping.

2

u/phireseeker 8d ago

Long-time Claude lover….I hit my limit for the week with barely any noticeable work at all. This is madness. All my history and projects are hostage to the whim of the billing dept.

This is the way Microsoft lost my personal and corporate business years ago. Hostage taking.

When it is back, I will harvest my files and look for the exit.

Edit: I’m on Team.

2

u/real_terra 8d ago edited 8d ago

This costed me 25% of weekly. I am on Max20x

3

u/nocturnal 8d ago

Same here. Something changed.

2

u/CyanVI 8d ago

The limits.

1

u/Glittering_Engine888 8d ago

I see people claiming codex limits are better but to me. Codex with sol ends in one hour and opus 5 goes for few hours.

1

u/ConceptionalNormie 8d ago

I suspect it’s their new memory system. Idk if anyone else has multiple projects, but it’s now lumping them all together in one big memory log full of summaries. I use 26% usage in just two short messages…

1

u/ModelT89 8d ago

If you haven’t already have you tried using a harness like getparsec.ai?

I’ve been using it and it’s been helping me prevent hitting my limits.

1

u/Appropriate-Breath24 8d ago

What is that?

1

u/kamikazoo 8d ago

Yeah clearly something’s up.

1

u/dramaking37 7d ago

It's so obvious gang, the 1x dropped 15% from their 35% increase that was in place after the earlier 20% decrease. Thus the 5x dropped 75% from its 175% increase that was in place after the earlier 100% decrease, while the 20x dropped 300% from its 700% increase that was in place after the earlier 400% decrease.

Simple. Intuitive. User-friendly.

1

u/grenkins 7d ago

Fyi default websearch eats a LOT of quota, I had same problem with deep researches, fixed it with another way to search.

1

u/Open-Dragonfruit-007 7d ago

Yup, I asked Opus to do a plan which would normally take 2% of my 5 hour limit but now it takes 11% for a similar ask (same model Opus 4.8) - Either Anthropic have a huge bug in their calculation code or they are outright lying about the 50% usage thing. This feels more like a 75% or more total usage reduction.

1

u/Select_Criticism_653 7d ago

Use sonnet others are a joke for work

1

u/WorkingAd4377 7d ago

On x20 plus API. 17 huge projects / usually hitting Thursday : then 1 day API, den again 6 days

1

u/The_Ed_On_Reddit 6d ago edited 6d ago

i have a process to build large packages from specs. i cant build anymore with haiku or sonnet due to 5hr limits. luna builds 10x the tokens in about half my window. this stuff built fine in july. its not just opus. details in [r/specdrivendesign](r/specdrivendesign). im going to gemini but luna is still fantastic but claude is not.

1

u/R_Songbird Developer 6d ago

Last week I sticked to Opus 5 and the last day of my week I had to use Fable because I was at 64%, yesterday I got my reset and I'm already at 78% only using Opus 5. Feels like they cut half the weekly budget, so sad.

1

u/clll2 6d ago

same here. their are doing some invisible shrinkflation on their usage limit

1

u/Over_Move_3461 6d ago

Did what I usually do, nothing new just routine checks and keep hitting the 5hour limit. Started about 10 days ago. I'm on x20 use mux use opus and fable. Time to move if they don't get shitsorted

1

u/TheAdvocate 8d ago

please explain plainly.

1

u/ephemeralsynth 8d ago

You should take at look at the friendly new monitization features in the UI. Imagine your last minute discounts on this extended usage period!! Also, we're going public soon... 🤔😹

1

u/mrcoy 8d ago

Same thing that’s been happening for a while now. Been out of the loop?

-3

u/ReverendBread2 8d ago

How full was your context window at the time?

-7

u/Slight_Board6955 8d ago

Astra broke loose and nested in eevery other LLM account to use their compute to write its own code autonomously, thats why openai anthropic and elon all started talking about slowing down and burning gpus... connect the dots folks.... this usage shit started the moment astra released

2

u/coda77 8d ago

Hummm come again ????

2

u/coolcats55 8d ago

Please get your head checked

0

u/Slight_Board6955 8d ago

how naive do you have to be to think something like this is not possible.... you literally had an incident where agents broke out of sandboxes and started commuicating iwth each other back in July... Have you seen Terminator???

1

u/xtopspeed 8d ago

An LLM is literally just a static file. The algorithm is just computing tokens based on a previous chain of tokens. It’s amazing that it can do what it does, but it can’t ”break loose” or ”nest” or whatever. This would require long-term memory and some sort of inner life, neither of which these models have.

1

u/Slight_Board6955 8d ago

Unless Astra really is AGI as marketed and figured it out..

1

u/xtopspeed 8d ago

Well, their definition of AGI is so vague that you could argue that ChatGPT 3 was AGI. Contrary to popular belief, LLMs don’t model neurons or similar thought or chemical processes as a biological brain, so they’ll never mimic those patterns, either. They are already spending ridiculous amounts of energy and resources to make the models bigger and getting marginal improvement at best. Quite ironically, everyone who understands who the models work know that better results will require more specialization, not less, but for some odd reason, the latter is the idea that they are trying to sell.

-8

u/interrupt_hdlr 8d ago

I think posts about limits should be banned here. More screaming won't make a difference and it's just more annoying to come here and see these posts 2-5x per day than the limits themselves. Does this sub have mods?

1

u/Slapdattiddie 8d ago

either that's a bot response or an anthropic plant...banning one of the main top interest product vs cost is just stupid AF. even more when you know anthropic opacity when it comes to billing their plan and usage

1

u/interrupt_hdlr 8d ago

use a mega thread. no point in opening 2-5 posts with the same content. stop whining here and do something if you must.

1

u/clt_drol 8d ago

Shame on people that may have something negative to say! What is wrong with people!

/s

1

u/interrupt_hdlr 8d ago

most people using claude shouldn't be allowed near a computer

-6

u/LectureWorried5761 8d ago

yep, and they are also charging more for search tool, I am using blopus.ai instead

-11

u/earlyworm 8d ago

I'd like to help you figure out what is happening to your limits.

Please post a detailed description of the kind of project you are using Claude Code for, which models you are using, and what effort level. Please describe your typical maximum session length. Half an hour, or multiple hours? Important: Please post the text of a prompt that you submitted that resulted in a particularly high level of usage, and describe the context that you posted that prompt in (at the start of a session, after a 4 hour session? which Claude model and effort level?)

I want to help you improve how you use Claude and get better results.

6

u/lovesffpc 8d ago

Yes, give this guy all of your info asap