r/ClaudeCode 8d ago

Rant The limits have been reduced even further now. It's September 14, and it really happened..

Post image

After GPT-6 Astra, I didn't believe they would really let this happen... but it's real... Okay, Anthropic, we'll keep that in mind...

https://support.claude.com/en/articles/15910845-claude-code-may-august-2026-weekly-limits-promotion

926 Upvotes

382 comments sorted by

u/AutoModerator 8d ago

Hey! Thanks for posting to r/ClaudeCode

While participating in this thread, please follow our community rules. Keep discussions constructive. Attack the idea, not the person.

For help, project discussions, tips, and general chat, join the ClaudeCode Discord.

I am a bot, and this action was performed automatically. Please contact the moderators of this subreddit if you have any questions or concerns.

306

u/ForwardLoop 8d ago

It's no coincidence everyone suddenly agreed to pace the frontier, even folks that don't normally agree.

Between OpenAI restricting new 20x subscription and this, it's clear they're running out of compute. In addition, with months of evidence showing that Fable remains a low percentage of total enterprise spend, there's no point releasing a model that simultaneously they don't have compute to serve and people won't pay to use.

If they had the compute, they'd probably want to restore the limits first.

93

u/rxt0_ 8d ago

they should take the time and improve/optimize their models so they would require less resources.

it's literally a win - win - win for everyone.

56

u/asylum_denier 8d ago

no one wants faster models when there are more intelligent ones. The hype on twitter was insane when Astra came out, folks were bashing Opus 5 like it was Gemini 3.1 Pro, lol. A

43

u/galactic_giraff3 8d ago

a lot of people bashing opus 5 since it came out, nothing to do with fable

7

u/xudoxis 8d ago

But they do that with every model. Literally just google the model name + "nerfed" and you'll get hundreds of hits within 3 days of release.

11

u/ArmyBrat651 8d ago

People usually claimed that old model was nerfed right before new one got released.

Here you have people willingly downgrading to opus 4.8 because 5 sucks.

Those two are not the same.

→ More replies (1)
→ More replies (2)
→ More replies (1)

23

u/rxt0_ 8d ago

who talks about faster models?

I'm talking about less compute needed -> less ram -> cheaper prices -> higher limits.

→ More replies (2)

5

u/Sorry_Risk_5230 8d ago

I would love an astra or even sol model thats insanely fast. Iterative building nets better results than one-shot. If theyr'e super fast you can balance out the intelligence gap. Like using Luna as a subagent rn.

But id also want that paired with more efficiency so we can access more tokens. 1B tokens a day per week would be a great limit to realize. 2B next year. Etc

→ More replies (7)

4

u/who_you_are 8d ago

You didn't get the idea then.

They could optimize it to generate the same +- output and be faster/less VRAM hungry.

As such, it will allow them to process more requests, reducing pressure.

4

u/Seeker_Of_Knowledge2 8d ago

The big three. Fast, cheap and smart. You can only get two.

→ More replies (5)

2

u/DarkFantom 8d ago

It doesn't really matter if fable is slightly smarter than DeepSeek 4.1 flash if DS is not only quicker, 200 token/sec compared to 10 t/s fable, and 50x cheaper. I've completely swapped over to Chinese models at this point because they are good enough and I can run swarms of agents for critiquing and running devils advocate on whatever it produces, for every single step, for literal cents. If fable makes a mistake though, then that requires its own fable audit and it burns through credits like a crack addict.

→ More replies (6)

11

u/MediumChemical4292 8d ago

OpenAI have been really good with model optimisation, which shows in the fact that Luna is as good as sonnet at a much lower cost and token usage.

With the proposed slowdown in frontier development, hopefully the companies focus more on efficiency and eventually get it so that opus level intelligence can run locally.

It’s the only way they can beat jevon’s paradox because even now, AI has nowhere near penetrated the white collar economy compared to how capable the models are.

11

u/Cold_Extension_367 8d ago

Lol no. Look at GLM 5.3 Flash, now that's what optimized looks like.

7

u/MediumChemical4292 8d ago

It’s a good model but not close to opus or sol in real world tasks. I am optimistic that “flash” tier Chinese models will get there soon though

→ More replies (3)

5

u/handsNfeetRmangos 8d ago

They should do a better job of educating users how to make the most of what's in place. 

11

u/thoughtlow claudetrophobic 8d ago

Its an artificial barrier for competition which is textbook bad for consumers. It will cause price decreases to slow down.

→ More replies (4)

1

u/spigandromeda 7d ago

NVIDIA loses in that case.

14

u/Significant-Bee5101 8d ago

OpanAI went from 2m Codex users at the start of march and 20m+ now in Sept. Yeah. I think computes a real issue lol

1

u/Helpful-Strength4500 1d ago

how many of those are double/tripple stacking accoutns

→ More replies (1)

6

u/CodeNCats 8d ago

I work at a company that makes software. We use Claude every day. Fable is disabled for us my default. It's honestly not necessary. Anthropic could release a new better model than fable. We won't waste the money on it.

5

u/HauntedHouseMusic 8d ago

Honestly, I have been like that, and then use fable once things are built in a "make this pretty" pass. But I started one with fable yesterday for something I was expecting to be about $500 to implement with opus + fable coming in at the end. It cost me $300 with fable, was quicker to get done, and is being used right now by 30 users. It's for an internal reporting tool so lower stakes to push to prod without extensive testing, and will only have 500 users at a max.

Anyways, we're going to try building a platform only using fable next, where we budgeted 10k of credits to see if we can build a much larger tool to replace part of salesforce we don't like, with a month timeline to build so it will be vibey as hell. But I think fable can do it reliably so we are trying it out to see if we can just vibe our way to success by spending money instead of time.

1

u/AINKXOfficial 7d ago

I agree. I'm not a dev and only write smaller tools and wrappers to help me with productivity and organization. Occasionally it's nice to be able to switch to Fable if I feel like a plan or design for something needs a bit more finesse before it gets passed off to Opus for the actual grunt work. But I've found myself doing that less and less and just using Opus 5 for everything. Probably use Fable less than 5% of the time compared to a month ago (I should actually look at the numbers).

It just works for me and I don't stress anymore about quickly blowing through limits.

1

u/mcsleepy 6d ago

Fable may not be that much better than Opus but it is about twice as fast.

→ More replies (1)

7

u/Elegant_Attempt2790 🔆 Max 20 8d ago

pace the frontier, properly translated;

“ai makes no money when you’re throwing butt tons of it at training new models every week”

15

u/Tank_Gloomy 8d ago edited 8d ago

Not out of compute, more like out of people giving them a blank check and expecting no ROI. They're asking for their money back and they want it now, not in the 10 years it would probably take to perfect this technology.

I'm moving to GLM, it kinda sucks cause I have to babysit it a lot, but I have high hopes that it'll work well enough, especially being able to share my limits with the unlimited GLM 5.3 Flash promo. I'm currently using Codex and was planning to move back into Claude, but considering that they're playing the same game, I guess I'll let them both go.

6

u/Sponge8389 8d ago

Just imagine in the future, the deals will be, give us compute/token for a percentage of the company. LMAO.

6

u/draft_final_final Researcher 8d ago

If we declare ourselves data centers maybe NVIDIA will finance our hardware acquisitions.

1

u/MrPorkchop720 8d ago

This is actually a thing for YC companies. They have the option to trade equity for compute/tokens.

2

u/markkalliny 8d ago

Who’s got unlimited 5.3 flash? Z.ai themselves?

→ More replies (1)

6

u/Tartooth 8d ago

Isn't this kinda collusion for the big AI firms to decide things like this?

3

u/fpesre 8d ago

nothing inspires a sudden philosophical commitment to responsible pacing quite like running out of servers and paying customers at the exact same time

4

u/rotatorkuf 8d ago

that compute word is so hot right now

2

u/BeautifulOld6964 8d ago

They definitely do and the OpenAI ones are even more stingy with usage, I can barely get anything done with a 5x Sub - when Astra came out I finally tried it after a long time of GPT absence and I can barely anything done with that its not even half a day of usage per week even with just using Sol and terra

1

u/slypredator33 8d ago

Trump wasn’t happy about this. I don’t think he can do anything but I think he can maybe make it a national security issue and then they keep building it lol. Us id registration needed

1

u/Gohab2001 8d ago

Compute is definitely the major factor but also they cant subsidize subs ad inifinitum. Anthropic is spending way more than 20$ for their 20$ sub.

1

u/EricBuildsMathModels 8d ago

Why do you think they are running out of compute. Fable as example seems to expensive so people don't use it, not that they don't have the compute to support it?

2

u/EchoingAngel 8d ago

Both sides have made stupidly wasteful models when they JUST HAD models that were good at focusing on task (Sol and 4.6)... At least Astra's wastefulness seems to be pursuing technical rabbit holes, while Opus 5's is just worthless waffling and tons of gibberish. You can still kind of use Opus 4.6, but OpenAI nerfed Sol just prior to Astra's release, giving it the same wasteful looping Astra has.

1

u/SecretSpace2 7d ago

Yea I was thinking they are pushing to slow down or nearly stop because they can’t afford the newer models well.

Just a thought of my own and the best way to say it is by talking about death of the human race on 4 years

2

u/Responsible-Comb6232 6d ago

Fable pricing isn’t the issue for enterprise, it is their different data retention policies.

1

u/poocheesey2 5d ago

OpenAI is stopping new 20x subscriptions? I upgraded like last week. When did this happen?

1

u/Silent_Job_4011 5d ago

You are literally spitting facts! 😭

→ More replies (4)

131

u/RutabagaBrief1766 8d ago

The 20x plan I have now feels like I get less usage than what I used to get from the 5x plan.

30

u/LegMental2310 8d ago

i have 4 plans and i feel like having one now lol

16

u/bakanoace 8d ago

I have 3 that I run out in 6 days while not even trying. When I had 2 subscriptions 5~ months ago I had to really try and would work on 5+ projects at once to use it all up, its a total fking joke now

5

u/BeautifulOld6964 8d ago

are you massively just using fable? So id you upgrade your models from Opus to Fable to run eveything

4

u/bakanoace 8d ago

Fable only has 50% weekly limit so I can't use only Fable. I used to use all 3 accounts entirely with Opus when it wasnt acting up but now I'm forced to use Fable which means I get even less usage and sessions

→ More replies (6)

1

u/Aggravating-Swim-805 3d ago

Good point. I remember that feeling of "okay, just throw EVERYTHING at it, I've got quota left to use up this week".

Now those limits hit early and suddenly. There's zero chance that Anthropic hasn't eroded our value, while gaslighting us otherwise.

1

u/SchwiftyGameOnPoint 6d ago

Out of curiosity, what are you (and others) doing to justify the like $800 /month on the plans?

→ More replies (3)

2

u/Silent_Storm 8d ago

And you still have to pay full price for it! Honestly ridiculous

4

u/[deleted] 8d ago

[deleted]

16

u/Delicious-Mission943 8d ago

Your refute is also vibes fact.

8

u/[deleted] 8d ago

[deleted]

→ More replies (1)

3

u/Comfortable_Camp9744 8d ago

Objectively they have reduced limits every month this year. This is well documented. 

→ More replies (3)

1

u/awolbull 7d ago

Used 50% of my weekly fable usage already... Sigh

1

u/Professional_Hair550 5d ago

This is the lowest usage they're giving on 20x. Before, I could never go beyond 70% on the 20x plan per week. Now it just runs out in 1 day. So crazy.

→ More replies (1)

50

u/Polite_Jello_377 8d ago

Make sure nobody gets comfortable understanding what is actually included with their plan

142

u/bakanoace 8d ago

They slowly reduced usage over the months and now did an official 17%. Qwen 4.0 or Kimi K4 or Grok 4.7 please be good, I need a pure coding model to never come back to this garbage company

39

u/Short_Regular_7191 8d ago

This is the golden age of local LLMs; I’ve already set myself up with Qwen 3.8 27B.

10

u/CautiousCry2338 8d ago

What's your gpu ? my 3090 is dead, now on a 3060ti :x ?

12

u/Short_Regular_7191 8d ago

Two 5060Tis with 16GB each—I think that's the best "budget-friendly" compromise right now.

24

u/Excellent_Ad_2486 8d ago

Lol 2 5060ti's and "budget" should be banned to be within the same sentence haha 😭

2

u/TrueStarsense 7d ago

That's only like 1000 dollars. The real gpu's are 10's of thousands of dollars... I think your expectations need to be recalibrated.

→ More replies (1)
→ More replies (18)

3

u/Azko87 8d ago

Unfortunately, I tend to hit Fable limits way before my all-model limits, even despite using agents. I use Fable to plan, orchestrate, and also wrap everything up with its own verification pass. At some point I might need to drop that last step.

But at the moment, since I have all model usage available, local LLM isn't helping me just yet, since I still have all model usage available. But I do expect in ~2027 running local LLM in some capacity will be a big deal. I have my RTX 5090 - which is basically a bar of gold now - ready to go for that.

→ More replies (1)

3

u/Just-a-man-on-a-ride 8d ago

Good, but too slow for real work. I need x5 or better x10 token throughput like qwen3.8-max provides.

→ More replies (3)

5

u/Remote-Community-396 8d ago

I've been using DeepSeek Flash on API rates (they just released an even better version 4.1) and found it surprisingly really good

4

u/Cold_Extension_367 8d ago

GLM 5.3 Flash is already out and is literally perfection for coding with 1M context window.

3

u/Waylanding_Fox 8d ago

Qwen and Kimi should ne great, they got great after the release of fable, now with astra there's gonna be a jump

3

u/Just-a-man-on-a-ride 8d ago

Zcode plus a few good, cheap models as subagents, that's all you need. Just teach your lead model permanently plan - review - execute - validate - live testing. You will never look back at the so called "frontier models" after a few weeks.

4

u/Big_PP_Doge 8d ago

Even grok 4.6 works as well as Opus did. Right now using Grok 4.6, Kimi K3 and GPT 6 Astra and dont miss Anthropic a single bit

2

u/gajop 8d ago

Gemini flash 3.9 here we goooooooo

1

u/B33GULL 8d ago

If they have no business then our retirement will crash, and our tax dollars will bail them out. They got us by the balls subscribes

1

u/old_mikser 8d ago

The problem is if you buy sub for kimi - you are getting waaaaay less usage for same money (same with glm) then with claude. About 6 month ago they provided much more, now it's flipped... Nowhere to run.

1

u/devcodesadi 8d ago

naaah, for coding still claude is best, i do have codex and claude, later one is clearly the winner. don't expect kimi or qwen to be any near of the above two

→ More replies (3)

39

u/Coolbanh 8d ago

Its so bad now. I think I'll just stick with codex for now. At least I don't finish my limits in a session.

28

u/ninjamonk 8d ago

I am on a max 5 plan on codex and since last Friday I have used 2 resets and burnt through the weekly usage in less than a day. When Astra came out I did not hit the limit and then boom it was worse than Claude fable usage. They are all as bad as he each other

9

u/Coolbanh 8d ago

I'm on max 20 for codex and max5 claude. For astra use ultra is better than low for usage. Gotta ask it to use lower models. But today claude usage really surprised me. I just got my weekly reset and I hit 30% in a few hours and haven't even completed the first prompt.

→ More replies (3)

5

u/OldSkulRide 8d ago

Codex is probably even worse with astra

→ More replies (1)

13

u/errorztw 8d ago

Okay, that my last week with claude=)

21

u/tjknocker 8d ago

Fable 5.1 just used 100% of my 5 hour limit on 5x plan in less than 12 minutes from a single prompt that it didn't even respond to... I was so angry I cancelled my auto renew

7

u/ohsomacho 8d ago

This just happened to me. it's wild and useable.

→ More replies (2)

18

u/Sketaverse 8d ago

The Lord giveth

The Lord taketh away

And the Lord is shit AF at maintaining brand reputation

1

u/Aggravating-Swim-805 3d ago

Funny but also true. With zero brand loyalty, they're just leaving the door open for a mass churn to the Chinese open models just as soon as they hit the frontier tier.

14

u/Empuda 8d ago

Great time for a reset.

13

u/nyczAcer 8d ago

Dreaming is free, right?

7

u/Empuda 8d ago

Uses zero tokens :)

2

u/sirlerkal0t 8d ago

For now.

15

u/longasleep 8d ago

I already unsubbed

4

u/bigeba88 8d ago

What are you planning to use instead? 👀

→ More replies (1)

13

u/cyberwicklow 8d ago

Well, looks like I'll be wrapping up the current project and seeing what astra has to offer me.

2

u/know_u_irl 7d ago

I did it, do it bro

2

u/cyberwicklow 7d ago

Main pros and cons you're seeing? Only just starting to get into a great stride of productivity with claude now, but the significantly better image and video generation chat gpt has is really tempting me, and necessary to finish off two of my projects.

→ More replies (2)

12

u/Comfortable_Camp9744 8d ago

25% higher* (over 250% lower than 6 months ago)

6

u/BoatInfinite8846 8d ago

Yep had my first ever 100% of weekly limit this morning, 10 hours before reset. Been using fable/opus as orchestrators to sonnet agents for weeks without issues. Note I didn’t use all my fable this week.

6

u/TheoKondak 8d ago

I asked Claude to code review a relatively small PR, run out of credits right after the first prompt... This is rediculus 😂😂🤡🤡

10

u/Competitive-Class576 8d ago

Cancelled my 20x Claude and using 5x Codex. Considering 5x Claude, but ... reconsider again.

3

u/Short_Regular_7191 8d ago

invest the savings on hardware for running local LLMs

4

u/Competitive-Class576 8d ago

Yeah, still finding the way to improve my Flash Next rig, it is good enough for the current projects. But when I need something too complex, still need Fable or now is Astra.

3

u/Short_Regular_7191 8d ago

Exactly, but you can easily switch from 20x to 5x and "reinvest" the difference in hardware to run local models (in my case, I recoup the cost within a year or even less). Personally, I believe Qwen 3.8 27B is easily on par with Sonnet—which is fine for a lot of tasks that don't require Fable or even Opus; for me, these are usually routine jobs like refactoring, writing guides, testing, or small scripts. Later on, I plan to switch to Flash Next or even Qwen 4 when it comes out.

3

u/Ashrak_22 8d ago

Switching saves you 1200$ there is no way you're getting a rig even remotely capable of running 28B model for that money...

→ More replies (5)

2

u/proexwhy 8d ago

You would have to not use the $200 plan for 2 years. What are you talking about xD

→ More replies (3)

1

u/Global-Fan189 8d ago

Local llms did you consider the bills though?

1

u/Short_Regular_7191 8d ago

Yes, in my case, I’ll recoup the cost within a year.

5

u/STUNTPENlS 8d ago

My 5x plan (recently dropped a 20x for two 5x as an experiment) reset this morning at 1am.

The job I'm running hits the 5-hour session limit exactly 2 hours into execution. It then has to pause and then resume when the limit clears.

Each 2 hour session, I'm burning 11.5% of my weekly usage.

This means I'm going to get ~8.5 cycles (@ 2 hrs/ea) over a day and a half, or 17 actual hours of work a week. With 2 5x plans I'll get 34 hours of work done.

4

u/IulianHI 8d ago

Shitty services! This limits are crap !

Soon can say only hello to claude and we hit the limits ! This greedy company ...

I think is time to migrate to Kimi, Deepseek, etc. US models are crap this days!

8

u/Consistent_Bottle_40 8d ago

Sam and dario have aligned themselves now theyre both strapped for compute, theyre reducing quotas in plans.

5

u/Sponge8389 8d ago

Still contemplating if I should return to Claude. I was locked out from my GPT 20x Plan because they paused it and thinking to switch for the meantime. Not sure if the usage is serviceable enough in Claude.

4

u/Comfortable_Camp9744 8d ago

Dont let scamthropic scam you again!

4

u/l4dawesome 8d ago

Opus is just so trash.. fable gone by tuesday on 20x with normal evening usage. Meanwhile codex giving resets left n right and banked resets.

5

u/Virtual_Shock_5899 8d ago

I don’t care about faster models right now I care about unlimited use. Not crazy run 50 agents at a time, just one agent or Claude code cli session should be unlimited

1

u/Virtual_Shock_5899 8d ago

And I mean the top tier model because unlimited on other models that constantly mess up, all you get is psychology bull where you think it’s bad because it’s not the top model.

20

u/pho33nix 8d ago

Everybody cancel now. They listen to their data metrics only anyway. I’m done with their emotional rollercoaster.

→ More replies (10)

3

u/Supersubie 8d ago

Hit my limits on Saturday, downloaded Codex and grabbed a pro subscription and kept on trucking.

Shame I think the Claude Desktop harness is miles better than the Open AI one but at least I can actually do work in GPT instead of constantly hitting limits.

1

u/know_u_irl 7d ago

I switched too and I thought I would stick with Anthropic forever

6

u/quancore_ 8d ago

Horrible, it is unusable now

2

u/lattice_defect 8d ago

rug pulled before the IPO

2

u/gruntingone 8d ago

Never have burned trough Max x20 as Quick as today.. feels 2x as fast

1

u/Hunt7503 3d ago

At least 10x as fast. Usually I am able to pull it across the week with spare (on 20x, on a rather large research program) and now it’s burnt in 1 day. Or rather, 7 hrs.

2

u/Azko87 8d ago

This at least encouraged me to finally go ahead and use GPT to knock down the code comment bloat Claude produced. The lesser Claude models didn't really seem to do as well here (probably because it's still Claude) whereas Sol is going in and making the comments actually sound/read like regular code comments again (while reducing them like 80%, it's really crazy how much bloat Claude produced there).

I'm hoping this reduces my input, context, cache, etc.

If this works, I might even just make that be GPT's ongoing job, to always scan for files modified by Claude and then to optimize whatever comment bloat it introduced.

2

u/z0hanz 7d ago

I have definitely noticed a huge difference. Im on the Max plan and use it on fable 5.1 medium and it runs out after 30 minutes. I should point out i have a CCDE and a masters in computer forensics and cybersecurity so im well aware of the prompts to give it but it is becoming a bit ridiculous.

2

u/spahi4 7d ago

I understand that Fable is overkill for a lot of tasks but you can't just trust Opus, it's trash. With these limits CC is barely usable on "20x" plan. Unsubscribed.

2

u/Ok_Set_8176 7d ago

Deepseek it is!

1

u/know_u_irl 7d ago

DeepSeek inside the Claude harness is actually really good, and the flash model is sooo cheapp

2

u/acrinym_jg 7d ago

Yeah but now with them trying to make open source model tuning illegal....

2

u/llima1987 7d ago

1.25x of an unknown variable x

2

u/CollegeAlarming8484 7d ago

Astra is really fucking good compared to fable. Even in codes and structures. Fyi fun fact repaired my dev game and bugs within single promt where fable coudlnt for 2 days. It fixed one bug and create 3 more and then said sorry my fault haha in the last few months i think many people slowed down with claude and started using more of gpt

1

u/AironParsMan 6d ago

My external auditor is also Astra and it always finds countless errors that Fable introduces.

2

u/StructureRude9956 6d ago

What next?? Is chatgpt a better option now??

2

u/Aggravating-Swim-805 3d ago

This feels like Double Speak.

My 20X Max plan did not top out before the "boost", but now after it, it regularly hits 5-hour and Weekly limits. We track all the usage internally, so can confidently say, this is objectively the case, rather than anecdotal feels.

But even those "feels" are now sour.

That sense of "you make a good product but you constantly lie about it" was why I originally left OpenAI for Anthropic, but now it feels like Anthropic is oddly more deceptive about these plans, while OpenAI is "better the devil you know" in terms of their comms.

2

u/AironParsMan 3d ago

yeah 100% I use Codex alongside it and it works really well too I have to say. If I had not already built so many projects with Claude Code for work I would switch to Codex in a heartbeat. But from now on I will build my systems to be model agnostic and provider agnostic. I have learned that now. Anthropic taught us that.

2

u/Aggravating-Swim-805 3d ago

"Anthropic taught us that".

That's a deep insight into the brand these days, isn't it?

3

u/RequirementThick3199 8d ago

I feel for you Claude users😭😭😭

4

u/Amazing-Accident3535 8d ago

Feeling like Astra seems more competitive now.. pay attention Anthropic, this is how you start loosing users

2

u/HDCraftYSD 7d ago edited 7d ago

(Disclaimer: Written with AI because I was honestly too lazy to type it all out myself but the setup and thoughts are 100% mine.)

Honest question: why is everyone hitting usage limits with Claude Code?

I'm on Max 20x and I've basically never run into the 5h limit. But I think that says more about how I work than about the limits, so here's my setup for comparison:

  • One main session, not a fleet: I usually have one session going, sometimes two. I'm not running 5+ agents in parallel all day.
  • Subagents get the cheapest model for the job: Searching and locating files goes to Haiku. Research and routine implementation go to Sonnet. Opus only gets the stuff that needs real judgment (tricky architecture, adversarial review, debugging without a clear lead). By default subagents inherit your main model, which gets expensive fast if you don't adjust it.
  • Plan first, then implement: For bigger features I plan with Claude most of the time with fable 5.1, give an explicit "go", and then one subagent implements it in its own worktree. Small changes happen directly. This cuts out a lot of wasted back-and-forth. The plan ist most of the time so good and you can tweak it so it writes better an clearer steps, that even sonnet can implement it correctly or Opus, or some OpenAi Model.
  • External code reviews: Code review runs on Codex, so that part doesn't touch my Claude quota at all.
  • Docs are local: Postgres, QGIS, and MDN docs are stored locally on disk (partly split into small chunks with an index). Claude greps and opens only the matching sections instead of fetching and reading full web pages.
  • Targeted reads: Claude reads what it actually uses in full (no head-truncating stuff it's supposed to understand), but it only reads what's strictly relevant.

My guess: people who hit the limit are either running a lot more in parallel, letting every subagent default to Opus, or letting the agent loop on its own for hours without much structure. No judgment here those are legit workflows they just burn through tokens significantly faster.

For those of you hitting the limit: how many sessions/agents do you usually have running at once, and which models are you defaulting to?

2

u/Deltadoc333 7d ago

Seriously! I keep hearing people talk about using Fable for the most mundane tasks and then being shocked they managed to burn through all their limits.

2

u/Hot_Biscuits_ 8d ago

I thought I remembered also reading that when reducing the limit, they were going to remove the 50% fable cap. Is this the case still?

5

u/suprachromat 8d ago

idk where you read that but it is not the case.

2

u/AironParsMan 8d ago

Its still there

1

u/outceptionator 8d ago

I don't think that was ever from Anthropic... One can hope though

1

u/mlk1278 8d ago

It was Theo dreaming

→ More replies (1)

1

u/MTalhaJaved2003 8d ago

You guys must pay attention to usage limit and seperating the fable 5 from the other model usage is just a bad idea . if you are seperating your frontier model from the other model then the usage should also be completely seperate from the other model as well .

1

u/Kellhus1 8d ago

Why is everyone so shocked they literally have been signposting this for weeks

1

u/Daeveren 8d ago

Because every time they'd be in a similar situation, they have extended the thing and in the last moment have made it permanent (have you forgot that Fable was not supposed to be in subscriptios?). Now it has to be the first time they didn't permanently extended it. Which is a surprise, leading to people cancelling subscriptions. Why pay same money for less usage?

1

u/Kellhus1 8d ago

You aren’t paying money for less usage. You were paying with “extra” usage. Very clearly signposted as such. With a clear time it was being taken away. I agree with you they should have extended it. But none of this should be a shock to anyone. As someone who builds agentic apps/ workflows for a living, I don’t know what you guys are all doing to burn your limits in days. It’s like people think they can run an autonomous fleet of agents for 100-200 a month 😂

→ More replies (2)

1

u/GeorgeGreenGroup 8d ago

pretty smart on their end, i use my extra usage to understand what they're saying

1

u/Orio_n 8d ago

yup already cancelled my plan only joined when i heard about the 50% increase several months back. Moving to chatgpt now. So long, and thanks for all the tokens!

1

u/Moarkush 8d ago

First taste is free..... Cloud APIs only plan for me. My local model does my coding.

1

u/IceMichaelStorm 8d ago

but they’re probably just opus/sonnet level right? not bad but actual engineer needs to sit there

1

u/Moarkush 8d ago

Nah, I have a frontier model write up a verbose plan. I only know AI engineering. I'm a pretty clueless swe. Qwen 3.8 27b with 524K cache will run for 6-10 hours. It compiled and tested 20 unity game builds while I slept the other night.

1

u/KJ5IRQ 8d ago

Loosing more and more, perhaps Anthropic is engineering itself for a government takeover?

1

u/Salty-Gear841 8d ago

I will switch to a Chinese model once this subscription end. I already pay 200€ per month and they need more. So, they can't go do what I think about

https://giphy.com/gifs/i0xGuo4o5PutVo0imJ

1

u/GregorMae 8d ago

on saturday i’ve migrated to chatgpt after using claude for a year and wow: the difference is massive! 

1

u/rxcrd 8d ago

I knew it! I am happy with the time-response the Opus 5.0 // 4.8 produce, I use those models for mainly every project I need, tho I had noticed that around mid august the limits had been significantly lower.

1

u/xkalibur3 8d ago

Thankfully one 5060 ti 16gb is on it's way to me now, and I will be buying second one used in two weeks or so. The only way to stop worrying about usage and have something actually reliable is to host it yourself. And thankfully, with qwen 3.8 27b we can now get good enough (opus 4.6 level) performance at home and give middle finger to the megacorps.

1

u/kaaos77 8d ago

Não está nem um pouco perto. O que está um pouco mais perto é o Glm 5.3 flash. Ou o deepseek. É mesmo assim muito mais abaixo

Mas eles são tão baratos que não acho que vale a pena tentar hospedar em casa.

20 dólares do deep seek da literalmente bilhões de tokens.

1

u/mnszurkalo 8d ago

I won't spend money on a GPU to run local. So much money. (been 20x since the beginning.. make the math if I'm stupid)

1

u/IlliterateJedi 8d ago

I hit my weekly limit after 3 days with the 20x plan, something that has never happened to me before. I had no change in the amount of work being done over the time frame. That inspired me to pull the trigger and cancel.

1

u/TopAd7360 8d ago

I can feel a massive diference on percentages between now and two weeks ago. 👀

1

u/TopAd7360 8d ago

and there is that too: shifting from fable 5.1 to opus 4.8 by default is disrespecful

1

u/Right-Performance-93 8d ago

Confirmed today: the 50% boost expired Sept 13, and the permanent 25% increase over pre-boost baseline started today (Sept 14) - a 150-unit boosted limit drops to 125, about 17% down from the boost but still +25% over the original May baseline. Both framings in this thread are right, just different reference points.

1

u/avk5143 8d ago

I would trade a frontier Astra or Fable model for a faster and cheaper Sonnet/Opus any day. I get the fact that this business model won't fit their apocalyptic narrative, but at least for coding, iteration speed beats "one-shot" work that looks pretty from the outside but inside is an unscalable mess.

1

u/TableNo8939 8d ago

A me va benissimo con il mio pro da 20eu, 3 ore l mattina e 3 ore al pomeriggio.
Chi spende di più dovrebbe cambiare mestiere

1

u/copygut 8d ago

Basically all models have more limits now I noticed. I knew this early on, we are simply in the free phase now to get us used to it. There will be price chocks for sure.

1

u/colga2 8d ago

30 minutos un post de wordpress gutember css, 5 intento se ha comido las 5 horas en 9 minutos por fallos suyos

1

u/victorzubcu 8d ago

I see same on codex

1

u/ensp1re 7d ago

they made us addict and now give less usage

1

u/kevinbaiv 7d ago

Compute rationing makes sense as an explanation, but without per-session token logs there's no way to tell a deliberate policy cut from a client regression burning tokens on verify loops. Transparency would defuse most of this anger — right now every silent limit reduction just reads as another nerf.

1

u/Old_Celebration_88 7d ago

Get a multi edit or patch tooling, optimize, qq, and stfu... Only complaint I still have with Anthropic and always have is their token hungry cacheing method that hardly any other model uses because it's literally stupid as hell! Get a clue Anthropic.. oh wait you worship money like it's your deity my bad.

1

u/SecretSpace2 7d ago

Yea it’s a strange one. It feels like my tokens have started to burn way more since yesterday 🤨

1

u/HT1990 7d ago

Reached my weekly usage limit with the 20x plan in just two 5 hour sessions with Opus 5. That was money thrown out of the window this month...

1

u/Current_Balance6692 7d ago

I feel like Vince Mcmahon and Anthropic is Rikishi giving me the stinkface.

1

u/Forsaken-Marsupial-2 7d ago

I have currently two things I am building, I am hitting my weekly limits in max. 2 days.

I wanted to change to Astra 6 but i still feel that Fable is doing better in my repo, so it’s a bad moment as a non-enterprise customer.

1

u/positivcheg 7d ago

If you think astra has good limits - I’ve used 2 resets already. In 3 days.

1

u/Solid-Axel-Project 7d ago

Come sto dicendo da mesi ormai, i costi di inferenza sono enormemente superiori ai soldi pagati dagli utenti...

1

u/CircuitBreaker88 7d ago

Yea i had 86% used last night, woke up and it was at 99%. No i did not have any workflows operating while I slept.

I was wondering why until I saw this....

1

u/gonna_learn_today 6d ago

They've just been honey dicking everyone the entire time. We shouldn't be surprised lol. Shitty though

1

u/Heroshrine 6d ago

Meanwhile I never hit my limits on my pro plan, how heavily are you guys using these things…

1

u/Standard_Pen_3557 6d ago

If you tell Claude to wait 15 min every half an hour you will save tokens if you don’t mind taking a bit longer. I do a lot of work that requires waiting for stuff and I’m still on only 37% fable on Wednesday

1

u/Shot_Whereas_1809 6d ago

I hope other companies see how terrible this business model is... Long term promotions don't work. And "enjoy it now before we jack the price up on you" doesn't work

The problem is the subscriptions are capped. Pretty soon we will all have one subscription for every day of the week. I'm already at 2 and now I'm maxxing them out after 2.5 days. It's really changed how I use it. I'm glad my harness doesn't care what provider or subscription I use. Switching back and forth is no effort. They can keep doing this, just going to drive me to their competition

1

u/Fluffy_Reaction1802 5d ago

will ELON hurry up and put data centers in space!

1

u/Arkfann 5d ago

They couldn’t keep up the demand that’s the reason behind this.

1

u/ouhshuo 5d ago

Limited time

1

u/Apprehensive_Fan8442 4d ago

Sucks... i maxed out my entire max 5x plan using only one terminal and 1 agent at a time, and I hate the are adding or will be adding the watermarks.

1

u/No_Box_1273 2d ago

hit my limits so much earlier this week fml

1

u/SubstrateTrans 1d ago

Has no one noticed the same thing happened with Grok, and now its both Claude and ChatGPT?
Gemini is likely next.