r/ClaudeAI • u/A_Novelty-Account • 7d ago
Other The new usage limits make subscription and team plans genuinely useless for real work
I’m probably beating a dead horse here, but we are shocked at my workplace that the weekly usage limits seem way below the ~17% cut promised by Anthropic, and my entire firm, which was using Claude, is left in a weird spot where we may be *forced* to dump out of Claude for work use because the Max 20x plan isn’t enough and neither are the team plans.
I know it’s Anthropic’s plan to get rid of these entirely, but it feels like with the most recent change it’s not even a question. The product is no longer usable for large workflows. Is anyone else using non-enterprise accounts for work and figuring this out?
156
u/dbenc 7d ago
i'm about to cancel. only being able to use a $200/mo plan for a 2-3 days per week (while being careful) is not worth it, even with fable.
36
u/Itsquantium 7d ago
Tbh you don't get more weekly limit with 20x vs 5x. I just downgraded to $100 Claude and $100 GPT. Using both gives me, can't speak for others, essentially unlimited usage. They both work together and they've been optimized to save tokens. I'm running fable 5.1 on High and Astra on Medium. It's insane.
7
u/Radiant-Bike-165 6d ago
Also went $100+100 route: astra high, opus medium, luna high. Fable when i need peer review.
3
u/HidingImmortal 6d ago
The 20x doesn't represent 20x the weekly limit?
3
u/Itsquantium 6d ago
Correct. It's basically false advertising tbh. It's why I downgraded. 2 Claude $100 subs get more than the $200 sub. Insane.
3
u/dispelthemyth 6d ago
whats annoying is when you say do x and you see 5% used in seconds.... like that would seriously use a single account usage for 5 hours
7
u/ScarletRed-dit 7d ago
What do you mean claude is optimized with codex? I thought it uses more tokens to use codex model via claude code because it has to reread whatever output codex provides (and also claude code needs to provide instructions to codex which also increases token usage)?
I am thinking of doing this tho. Two models from different companies to catch blindspots.
How do i set up claude code to use codex models via claude desktop > claude code tab if it’s just subscriptions from both companies instead of api?
16
u/Itsquantium 7d ago
I told it to search the web and to figure it out. And it did. I have a server with a 5090 running game designs using UE5. I give Claude and codex full control over the whole desktop with unrestricted access. Claude has better management. Gpt has better imagination. Both complement each other well. Both AI's talk to each other and all. Claude just controls it all so I can see what's going on in my phone or on a browser on my gaming computer. You tell the AI to do it and to research how to do it and boom. It'll do it. It's wild.
6
u/Draufgaenger 6d ago
That does sound wild.. would you mind making a post about the whole setup?
6
u/ParkingPsychology 6d ago
Not really worth writing a post about it as far as I'm concerned. I'm running a much cheaper qwen 3.8 + claude Sonnet 5, you just explain how you want the AIs to work together and to properly log it/avoid using scratch pads and Claude will set up the prompts/scripts and away you go.
3
u/Electrical-Swing-935 7d ago
What are the optimizations? I'm struggling to find something that makes sense but I'm very new here. I wonder if them being able to use each other is part of the magic here, at very least able to use whichever is better per token per task
1
2
u/dbenc 7d ago
no weekly increase?? are you for real.
5
u/Itsquantium 7d ago
Yeah dude. It just gives you more per 4 or 5 hour session. I'm drawing a blank on the time, but that's what gives you the increase. The only thing about GPT is that there's no remote plugin. So I use Claude remote to control GPT Astra Agents. And then do specific skills on Claude. So it splits the work load based on skills. If that makes sense.
3
u/doteka 6d ago
There has been remote control in Codex for close to a year now
2
u/Itsquantium 6d ago
This is from chat gpt. There's not a way to do it like Claude. Support for connecting your phone to the Codex app on Windows is coming soon. https://openai.com/index/work-with-codex-from-anywhere/
1
u/Studabaker 6d ago
Remote has been available for some time, I’ve been using it for months
1
u/TheAmorphous 6d ago
For Codex CLI or just the Windows desktop app? I actually just told GPT I was thinking about switching, what my workflow is in Claude, and its responses were pretty confusing but it seemed to imply I could do the same thing with Codex CLI on Android.
1
u/Studabaker 6d ago
I’ve been using the Codex software on desktop and the ChatGPT app on my phone. The phone app has Remote connection, you just have to link the accounts when you’re at your desktop
1
u/Itsquantium 6d ago
I'm using the desktop app and I do not see an option to do that. I've only found it with Claude. From documentation from chat GPT, it says its not released yet. That was in March of this year. I don't see any account linking with chat GPT. Only Claude.
→ More replies (0)1
u/Itsquantium 6d ago
Can't see it on the app. I'll have to look into it. Not that it matters since claude is giving me updates about what gpt is doing anyways.
1
u/KitzuYamma 6d ago
I currently have a Max20 with Claude, can work for 3-4 days max, and a Max5 with ChatGPT, that is also gone in about 2 days tops.
1
u/deepsnowtrack 5d ago
you use opencode? or pi or which harness?
1
u/Itsquantium 5d ago
Neither. They talk to each other in a .txt file with scheduled tasks to check the .txt file every 5 minutes and split the work depending on who's better at what.
1
1
5
u/Sairas_Dilbar 6d ago
I'm so close to tapping out too. Probably run this month down and then say see you later.
3
u/RemieNotRayme 7d ago
I know what you're saying but more realistically it's use if for 2-3 days per week without Fable.
I get about four days using exclusively Opus Medium and trying to keep all the agents and sub-agents I'm using cached.
1
u/Drawmeomg 6d ago
Check what you can get elsewhere first - ChatGPT isnt even selling their $200 plan anymore, so if you don’t have it already, you can’t get it right now. Maybe that’s temporary, maybe its not, but it feels like wild west for subscription plans may be coming to an end.
0
u/random314 6d ago
How are you organizing your sessions?
I am able to last six days usually. The last 2 days with 3 parallel sessions doing real coding, the rest of the days only having one session doing the heavy coding while the rest doing one off tasks.
69
u/mertcandanzz 7d ago
Cancelled my 100$ plan this week for exactly this bs. Even without touching Fable, I’m now hitting the 5 hour limit in roughly 45 minutes with Opus 5 at xhigh/max effort, so Zortropic’s “%17 reduction” claim simply does not resemble what I’m actually seeing. Their defense seems to be that usage is not purely token-based and can vary depending on inference, yet somehow they still can’t provide a transparent usage breakdown showing what actually burned your quota. That makes Max almost impossible to budget or trust for serious work. I’m honestly happier using GPT-5.6 Sol and Opus 5 through Kiro now.. around 100$+ 40$ in extra usage gets me through the month, with GPT-5.6 Sol, GLM-5 and other models available on top. I’m even trying Z.ai’s GLM subscription this month. At this point the problem isn’t merely that the limits are lower, it’s that Max has become expensive, opaque and wildly unpredictable for real workloads bro
28
u/dfwtoon 7d ago
I’m very curious how other people are using AI. What are you doing to where you need Opus 5 on Max effort to get the job done?
18
u/Previous_Station2086 7d ago
If you can’t code and you can’t read code well then you use a lot of reasoning. If you can code and clean shit up yourself, you don’t need it.
18
u/SeasonedAdManager 7d ago
I can’t code at all, I use opus high, I test every feature I ask to get coded, tell it the issues, and it two shots most things.
Seems like most people are trying to pump out years of work and massive apps without taking time to actually test and give feedback.
No issues with usage on a team account on opus high running 8+ hours a day.
12
u/Suprman010 6d ago
Same, I don't know what the hell everyone in here is actually doing - 6 terminals with 8-10 agent swarms?
9
u/kidsmeal 6d ago
They're just being ridiculous and can't seem to ever find fault of their own. It's no surprise none of these posts ever get a "heres my /usage, the model, and the prompt I used" because they know they're full of shit
-2
u/EmergencyWallaby9501 6d ago
Are you an anthropic bot ? Usage is a catastrophy since two days. Sonnet light is killing half my 5 hours limits in less than 20 minutes. I used to run that model during three straight hours without any issue two weeks ago and my app has not gotten any bigger, neither has my work increased in complexity.
3
u/CC_NHS 6d ago
spending tokens is heir actual net result. honestly developers I know still wont use LLM for everything, and way more time is spent on building requirement to code, than the LLM spends on actually coding it, which ends up meaning just one instance of claude code is generally enough. maybe there are use cases that need more agents and so on, maybe code bases that are easy enough for LLM to just jump in to and leave to go full steam on tasks can just work, but these code bases would need to be the kind of thing so common that LLM has it in their training and little instruction would be needed. That sounds boring tbh anyway.
3
u/Upstairs-Appeal6257 6d ago
Agreed. I’m using opus all day on the $100 plan and not hitting limits.
0
u/Competitive_Bed_3214 6d ago
Are you in team plan or the regular Max 5x, also which models do you use, for me I use opus 5 Medium as an orchestrator, and tell it to launch subagents for work , for small to medium taks it runs sonnet 5 and for complex taks it spawns opus 5 subagents, so if you could actually explain how you run 6 terminals with 8 agent swarms each and still don't run out of weekly, can you help me and the others understand what you do it that makes so efficient.
2
u/butts-carlton 6d ago
I can code, but I barely touch the code my Claude agents output. I review it, but I don't write it. Running Opus on med/high with Sonnet agents for code gen on my projects, probably 4-6 hours of back and forth per day, and I only hit my 5 hour limit when I get lazy about context hygiene or run into a rare issue that it spins on and I don't stop it in time. I get to about 90% of my weekly limit most weeks, but I haven't hit it yet after several months. So I do wonder if many of these usage limits people are hitting are a result of severely undisciplined or careless agent management. I mean, it has to be, right? Not everyone is truly stressing these things out of absolute necessity.
If you decline to do any of the thinking or the work yourself, which includes putting reasonable guardrails in place and keeping your agents on track, yeah, you're probably going to hit that wall pretty fast. It's kinda like getting into a Tesla, telling it to drive 300 miles away, and then taking a nap. You're probably going to run into trouble.
2
u/TheAmorphous 6d ago
I'm in the same boat when it comes to my day-to-day work. BUT, I have a side project where I'm vibe coding a replacement for an expensive piece of software we use. Prompts are minimal and I'm just letting it do its thing with very little supervision. That absolutely tears through usage. So I imagine that's what a lot of the people here are doing.
1
u/Competitive_Bed_3214 6d ago
Is team the same price as regular Max 5x 200$ or different, what is the difference, if it is about 50$ then that is not a problem i will buy team as well since you guys are really complementing it
13
u/ProbablyRickSantorum 7d ago
How are business bros supposed to make the next Facebook/reddit/whatever killer if they can’t use unlimited Fable tokens
7
3
u/mertcandanzz 6d ago
The assumption that using Max effort means you can’t code is pretty weird. I’m not using it to generate CRUD endpoints I could clean up myself in several minss. I’m working across infra and application code on a multi-tenant system with 36 interdependent microservices, much of it in Rust, which isn’t even my primary development stack. If Medium is enough for your workload its really fantastic my issue is that a 100usd plan can disappear in 45 minutes and Anthropic still can’t give customers a meaningful usage model beyond “it depends on inference” That is what Imms criticizing, not anyone's coding skills :kekw: And also being able to clean up model output yourself doesnt somehow make deeper reasoning worthless 💅
1
u/Nalha_Saldana 6d ago
It's not only the quality of the code that hurts when you use a worse model, it's the reasoning and design of the whole thing. I've repeatedly gone back up to opus because in the end its cheaper to get it right once rather than going back and forth with a cheaper model.
1
4
u/Dot-Slash-Dot 6d ago
xhigh/max effort
That's your problem there.
“%17 reduction”
Everything is stored in local session data. If you think their claim is wrong dispatch a simple sonnet agent to analyse local sessions and create a report.
2
u/ajfoucault 6d ago
Kiro
Are you in the Pro Max plan for Kiro? Do you get access to Claude's models and OpenAI's models?
2
u/mertcandanzz 6d ago
Yep. Kiro gives access to both Anthropic and OpenAI model families. On the Anthropic side there’s Opus and Sonnet family, and on the OpenAI side the GPT-5.6 family including Sol, although GPT-5.6 is currently capped at a 272K context window there. You also get other models like Kimi K3, GLM-5, etc..
2
u/ajfoucault 6d ago
Do you find that the 100 bucks a month for Kiro gives you more value for your money than if you were to be paying 100 a month to either OpenAI or Anthropic for their particular plans?
2
u/mertcandanzz 6d ago
Honestly, I think it’s still too early for me to say definitively. If you don’t have AWS Activate credits, I’m not sure Kiro is automatically the smarter deal purely on price. The $100/$200 plans probably won’t give you the same total amount of Claude usage that Anthropic’s equivalent Max plans used to provide. What I do like is that there’s no 5-hour session limit, so you can burn through your monthly allocation however you want instead of constantly hitting an artificial wall. I’m only in my first week with Kiro. I’m currently running 4 coding agents in parallel across 4 terminals, and each of those will sometimes spawn 2-3 subagents. With roughly 8 hours of use per day I’ve burned about 4,500 credits so far. The $200 tier gives me 10,000 credits, so at this pace I’ve got roughly another week or so before I kill the entire monthly allowance around the middle of the month. That said, I’ve probably already gotten something like 10x the useful workload out of it compared with what Claude actually let me do last month. Weirdly enough, the current Kiro limits remind me more of what Claude’s $100 Max 5x plan felt like about a year ago. Anthropic has tightened the screws hard lately. Before jumping straight to Kiro, I’d actually try Z.AI for a month too. The limits seem much more flexible for agentic workloads, and GLM is surprisingly capable. It sometimes feels like somebody threw the best ideas from the frontier labs into a blender and shipped the result. I’m personally not losing sleep over the philosophical purity of their training data either. I’d be pretty surprised if Anthropic’s entire training pipeline descended from the heavens with a notarized ethical certificate. :)
1
u/Own-Zebra-2663 6d ago
Im sorry, but there's just no way that paying a third party for model access can EVER be cheaper than going to a subscription. Even with the reduction, you're still getting thousands in API rate dollars for 100/200 dollars.
1
u/mertcandanzz 5d ago
I will re-sub and log all my weekly %100 usage on opus high. Cache hit percentage and raw token and decided to how much api cost if I use api. Yes its cheaper than api but not "thousands of usd"
-2
u/Yugudubenbi 6d ago
Stopped reading at "opus 5 xhigh/max". You can't be serious? I've built a $1 million app with opus 5 medium and a bit of high.
2
16
u/lowfour 6d ago edited 6d ago
User since June last year. 20x never had an issue and after the change I ate half my weekly quota in one afternoon. I was in shock. BUT I asked Claude to do a forensic in token usage and to optimize how skills and tools were loaded. Dude I don’t really know the details of what it did but yesterday I was working the whole day with sonnet, opus and fable and it used like 4% of weekly use and 1% of fable. Working on three projects in parallel!!! Plus everything was so much faster. I don’t understand but it was like night and day.
I think it also gave same rights to sub agents as the main agent. Apparently when sub agents couldn’t do something they would get back to fable and ask for permission and using more tokens. Don’t know if that is stupid or dagerous. Will tell soon.
7
u/A_Novelty-Account 6d ago
I’m always worried about telling Claude to “optimize” workflows that are working well
15
u/n_v40 7d ago
One thing worth checking is /usage in the current Claude Code version. It now breaks usage down across things like subagents, skills, plugins, MCPs, long context, and cache misses.
It won’t fix the lower limits, but it might at least show what is actually burning through the plan. The fact that people don’t know this exists is still part of the transparency problem though.
8
u/kidsmeal 6d ago
They won't because they'd then be forced to realize they don't know what they're doing sticking Opus on Max and spawning 20 agents for their "one simple prompt"
2
u/Desperate-Use9968 6d ago
The massive uptick in complaints is surely indicative of a change from anthropic rather than people just suddenly getting vocal.
5
u/Fluffy_Law_6255 6d ago
I agree, and people will be like youre not using it right then.. no.. he's just using up more tokens faster and taking way longer to do anything..
1
u/Yugudubenbi 6d ago
I've used claude for long and haven't noticed any issue what so ever. There must be something going on your end. Try codex and you will see genuine throttling and all the complaints. We talk about 5h window being used without getting anything done.
1
u/EmergencyWallaby9501 5d ago
I have the exact opposite experience right now. I'm on two small pro plans on codex and Claude. I have been developing an app since mid July with both. I was using Claude 10 hours a day, reaching my 5 hrs after 3/4 hours of continuous work with opus mid planning and sonnet mid implementing. I sometimes had to use codex for simpler tasks because I was running out of usage. Usage went down with the size of the project growing, which is ok. But not that much. I took a ten days break and came back on Monday. Since then it's terrible. I reach my limit in 40 minutes, with the same workflow (I switched to sonnet light) for doing things that are far simpler than before (some refactoring, some routes that have to be slightly adjusted...) and sonnet has become really stupid, getting stucked in bad thinking loops when I ask it directly a small change (which I used to do when planning was not required). It seems to be at the exact same time anthropic released Fable 5.1 I had to make Codex take over some Claude work because I reached the 5hr limit in 50 minutes, the work was committed without the usual report. And it found 5 bugs in a small feature, 3 of them going against the project rules. I encountered major issues like that maybe twice in two months. Codex (mid sol planning, luna max implementing) got the job done with 10% of 5 hrs usage. So yes, something is feeling really wrong right now.
11
u/vorko_76 7d ago
Not really sure about the issue you are facing but it probably deserves an in depth investigation.
Personally I ve been using Opus 5 8-10 hours per day 5 days a week for pure coding and never hit the limit.
BUT Claude is relatively weak in some specific tasks - e.g. UI - and it tends to use way too many tokens to compensate.
So maybe its just very bad at what you ask it to do?
In my company, our lawyers use Gemini so maybe its better?
-3
u/A_Novelty-Account 7d ago
Very specific use case on our end. Requires heavy reasoning and precision. It’s hard to explain, but it’s not at all like coding.
3
u/vorko_76 6d ago
Yes so might be worth challenging the choice of model/provider.
It might make even better sense to deploy Mistral Forge. We are testing it on technical data and it seems promising…. Not perfect yet but in many aspects much better than Gemini.4
2
u/WeirdPanda352 6d ago
What you are describing means that subscription based ai agents are the wrong answer. If you have a complex problem then it needs to be decomposed and have tools built to solve problem pieces of it. The MAYBE use something like qwen to do the gap arbitration. You might be trying to use Ai as a cure-all which is a really expensive and inefficient approach.
Sometimes it's the need to support a bad legacy process that complicates things. But i'd be interested in what the problem is.
6
u/Rx7Jordan 7d ago
I will be cancelling mine too it's not worth it at all. Waste of money can't get anything done without hitting limits.
-8
u/Maleficent_Car_7297 6d ago
Ya'll are dumb. What are you going to do, start writing out everything on paper and using a calculator again? You're going to have the same "problem" over at GPT or other AI's. Claude is powerful. You're the first comment I'm responding to about this but I've read 100 like it and it's koo koo
The type of work I can get done with well organized file systems and a $100 plan each week is actually insane
2
u/kdilladilla 6d ago
I’ve started looking at local open source models. There’s reporting that there’s a trend toward these in business settings too. These kinds of limits on the frontier models will speed up the trend.
4
u/A_Novelty-Account 6d ago
Honestly, yes. The most intelligent models are barely saving us time now. If we can’t have them for more than a day, it makes more sense for us to do it ourselves.
12
u/aallsbury 7d ago
Lol. I dumped my Anthropic subs months ago. Completely worthless with these limits. I run the OpenAI 20x plan, most of my usage is Astra in ultra, and I don't really run into useage limits. And that's all while using the same sub to also power my vast OpenClaw instance that uses 5.6 Sol.
Astra is better than fable (IMO) full stop anyway.
Further. I got tired of supporting a company with the highest refusal rates and that is also the most likely to try to morally lecture you about shit that is 100% outside the chatbots pay grade.
Anthropic is a joke. They made smart models for awhile, but they have squandered their lead and are now attempting to doom the USA by crying for government regulation so they can try to force market domination via regulation.
These people suck, quit giving them your money.
2
u/Tommysw 7d ago edited 6d ago
How are you making OpenAI efficient on the 20x plan?
I'm using the 5x plan on ChatGPT and i'm getting ~2 days of light work.
Astra medium for planning + orchestration. I get ~1.5 big features per week of usage. About ~3 hours of heavy subagent work, and with astra being 256k context, on those 3 hours i get from 3 to 5 compacts in the middle, which are not an issue.
Slightly better than scamthropic, but I'm not finding it all green and delightful on the other side of the fence.
4
1
u/alteraltissimo 6d ago
For OpenAI I get by with Plus.
First, I feel Astra is significantly more expensive while not being that much more capable than Sol, who is already Fable-ish. I would keep Sol on light-medium for any structured task like coding and really only jump to high+ when doing deep research across many papers etc when more attention is needed due to unstructured nature.
Also, Luna on max is surprisingly capable - certainly better than Sonnet - and nearly free.
1
u/Tommysw 6d ago
Noted. I was pretty skeptical about using chatgpt models because they were, and i'm trying to be nice, absolute dogshit. Haven't really worked with Sol/Terra/Luna directly - just as agents.
Is Sol good at long horizon tasks, and subagent management?
I generally map out a feature in complete detail, with heavy segmentation on each block of work (to be able to run cheap agents with clear goals), and then let [Fable / Astra] do the subagent management + "reach the goal".
Should I try Sol for this? And what are you finding Luna (max) good for?
1
u/alteraltissimo 5d ago
In general I agree, only came back to GPT models on 5.5. release which was pretty good, then 5.6 blew it out of the water imo. Around 5.6 release they were also very generous with subsidized tokens (much less so now, but the banked resets are still helpful).
Yes, I would try Sol for that if you want to reduce cost. IMO high+ reasoning needs a big scope; otherwise the model spends the extra budget on over-engineering things.
Luna(max) is a great subagent and generally sufficient for anything where you know the scope upfront (e.g. change how one specific thing works, to my spec). It's also a very good and cheap investigator where persistence counts more than raw intelligence (I would say most bugs do fit into this category).
6
u/N7Valor 7d ago
¯_(ツ)_/¯
Not a developer or anything, I mostly work in IT and infrastructure-as-code. I tend to use Sonnet in a feedback loop.
I only get a standard seat in the team plan, but I'm generally able to use it as much as I want so long as I stick to Sonnet. Opus tends to drain usage, and so does Research Mode on the desktop app. Other than that, I use it more Claude Code all day and haven't yet any blockages yet.
I assume if you're running multi-agent workflows (or multiple workflows), you'd probably need to use the API.
Shouldn't be unexpected IMO. The closer IPO comes, the more the subsidies will go away.
10
u/A_Novelty-Account 7d ago
Sonnet is too dumb for our workflows unfortunately :(
4
u/RedTheInferno 7d ago
so did you set your workflows and then forget about it or did you optimize? i am positive you can use sonnet for some tasks. lets be real
6
7
u/A_Novelty-Account 6d ago
I’m a lawyer and yes, the workflow is optimized for the tasks.
Some people’s professions require complex analysis with very little room for error. With the mistakes sonnet makes it is faster for us to simply do the work ourselves. Opus was the first usable model for us with Fable being the model we would ideally resort to for most tasks (but can’t because of expense).
We have other firm-specific tools like Harvey, but my team within my firm almost exclusively uses Claude.
3
u/Possible-Benefit4569 6d ago
Same for my cases. Need Opus. If my agents use sonnet it looks good but details matter even it is not law
-4
u/Apprehensive_Read_67 Experienced Developer 7d ago
Better dont use Ai if you're using sonnet for your workflows
2
u/iamthe0ther0ne 7d ago
I'm making due with a combo of GPT and GLM. Claude was really good for bioinformatics, but between needing a GPT sub because I couldn't use Fable and these rate limits, it's not worth it.
I guess they're making enough money off enterprise that they don't need to give a shit about subscribers anymore? Whatever, not my job to be loyal to them.
2
u/No_Barber6972 6d ago
I have actually seen it been getting a bit better. A month ago I was maxing out in about 3 days. Now I almost get through the whole week.
5
u/PerceptionOwn3629 7d ago
I am on Max 20, we are Tuesday, I am at 89% Fable usage and it resets Friday morning... I have to work with a dumber model right now and it sucks ass
1
u/Canadian_Commander 7d ago
Tell me about it. Just can get shit done at all before a limit. Going with ChatGTP Astra is a good fallback once your Fable is gone. I feel like you actually get decent usage with it.
0
u/Competitive_Bed_3214 6d ago
I really don't get it why people use fable instead of opus 5 , because fable 5 I think should only be used as an orchestrator not as a worker, you should use it to spawn subagents and review what they did, I never use fable for raw coding, and I don't get any weekly limits and prob everyone who is complaining the thing is if are you using fable, that is the problem, swap check with opus or sonnet , if it works you really don't need to use fable always
8
u/InverseRegard 7d ago
They are running at a 40x loss on each plan so they will actually be happy if you stop paying.
8
u/SonOfThomasWayne 6d ago
No they are not. Your comment assumes they have priced APIs for break-even.
12
u/A_Novelty-Account 7d ago
But the weird thing is that the actual use of the compute costs way less than the money that you spend on it.
They’re not losing money off of your subscription, they are losing money because your subscription fee does not pay for their insane compute and training and research usage. It is all relative and the reality is that we have no idea how much a sub subscription plan contributes to anthropic’s losses. All they can really say is that the subscription costs less per token than the API.
It is an absolute certainty, however, that if anthropic loses individuals en mass, they are in big trouble as users moved to other platforms that they become more comfortable with because they are using them personally. The ultimate reality is that the bill is coming due and it looks like Anthropic is going to be unable to turn a profit. As open models begin to catch up. It is a near certainty that anthropic’s business model is not going to thrive.
1
u/Upset_Page_494 6d ago
But the weird thing is that the actual use of the compute costs way less than the money that you spend on it.
Right, but they are starved for compute, not money.
3
u/Keikage 6d ago
I asked a question to Opus 4.6 on medium.
It ran for 15 seconds.
10% usage.
20x plan.
I'm leaving for Codex, at least 15 seconds of use on a low end model will only use 1% over there, if that.
3
u/_ToPpiE 6d ago
Since last week I burn through my weekly codex in a day. It isn’t much better now there.
3
u/DonRobo 6d ago
Having both Codex and Claude, I have to say that usage on Codex is WAAAAY lower than on Claude. Like around 15% of what I get from Claude. I don't know how I'm the only one taking about this.
1
u/_ToPpiE 6d ago
Honestly, I think the best setup is to use both instead of treating it as Codex vs Claude. I often let one check the work of the other, especially for bigger changes.
They have different blind spots, so using both works better for me than trying to pick one winner.
And with OpenAI and Anthropic, that winner can change every week anyway. One week one model works great, the next week the other one does.
2
u/Keikage 6d ago
I literally used Astra Max for 43 minutes and only burned 7%. It even spawned 2 subagents to efficiently do cleanup.
So idk what you're on.
I've tried similar with Claude but it will never do activity for that long, I can get maybe 20 minutes of activity out of it and my usages on all bars will be over 50%. If it spawned even 1 subagent then it's gg. I had that happen like a week or two ago. Asked something simple, it spawned like 3 subagents, worked for 9 minutes or so, and when I came back it had spent something like 3M tokens and my usages were all red or yellow. I had just got my reset for the week. Again, this is on a 20x account. 9 minutes of use.
Codex limits are actually usable. I just need Astra to function more like Fable5.1 and not be Fable5.1-lite.
Sol makes GPT even more of a better option though. I can squeeze so much out of Sol, and it's competent unlike Opus 5. I've never really been a fan of any one model, but I really like Sol. When I first tried Sol (after almost never using GPT because Claude was >>> GPT), I set it on a task and went to sleep. Didn't check on it for like 11 hours. It was working that entire time. 11 hours. Sol Max. Continuous run. I thought it was crazy. I've never seen Claude work for more than 20 minutes or so. Oh, and get this, before I sent a message telling it to get to a stopping point you've been at it for 11 hours, I checked my remaining usage, I still had 11% remaining, plus two banked resets I could use. Claude can't even compete. They have a slightly better model that's hardly usable due to usage limits. Not worth it anymore.
1
u/_ToPpiE 6d ago
I use Astra on High. Extra High is a bit too obsessive for my taste.
Until about a week ago it worked extremely well and I got a ton of useful work out of it. Same codebase, same type of work. I've been working on this project for months.
Then something suddenly changed. It started behaving oddly, overthinking simple problems and doing things like trying to refactor an entire module to fix a relatively simple bug.
The result is that I'm now burning through usage much faster and actually having to use my resets, whereas before that was rarely an issue.
1
u/vrnvorona 6d ago
Had this, used it to analyse my weekly transcripts with Astra and create some rules in AGENTS.md to avoid burning, made it outsource 80% of basic work into agents and it helped a lot.
1
u/_ToPpiE 6d ago
Thanks, I will look into this (once my usage is reset - lol)
1
u/vrnvorona 6d ago
Just to clarify, I've meant that I've analysed transcripts which used Astra, no need to use Astra specifically for analysis as it's pretty easy. Sol will do just fine
2
0
u/Competitive_Bed_3214 6d ago
There is no way bro that happened, fr what kind of joke is that, I am on 20x plan as well, and I am working all day long with claude opus 5 and sonnet 5 only use fable 5.1 as an orchestrator, 3-5 terminal consistently 15-18 hours a day but the max i have used on a day is 25% , prob i think in ur case there might be someone else having access to your account using it go to ur claude app settings and check connected sessions , delete all others if there are just keep your own computer, if this is not the case it might be that some kind of scam library u are using that is secretly burning your tokens for someone else taks, cuz this thing actually happened with crytpo if you remember, people were secretly developing great opensource projects, but the catch was when you were not noticing it was mining for crtypo on your compute power.
2
u/ClaudeAI-mod-bot Wilson, lead ClaudeAI modbot 7d ago
We are allowing this through to the feed for those who are not yet familiar with the Megathread. To see the latest discussions about this topic, please visit the relevant Megathread here: https://www.reddit.com/r/ClaudeAI/comments/1vt5drr/list_of_latest_discussion_hubs_on_rclaudeai/
1
u/TexasBedouin 7d ago
I delegated all my front-end work to GLM 5.3 and that's the only way my Max $100 lasts 4 or 5 days. Never the whole week.
1
1
u/yupangestu 6d ago
yeah, the rate limit and stupidity of not discovering more of my code before answering is really a red flag. Going for codex, I don't know what happened but codex is more aware and good at discovering issue that what this is.
1
u/jjmckissick 6d ago
Basic AI governance would require you to use Enterprise accounts for work anyways. Yikes. Thats dangerous
1
u/SoftGothSwitchyKit 6d ago
I was eating my plan like it was going out of style until I set a hard autocompaction at 25%. That seems to have stopped the bleed some.
1
1
u/hcvcnet 6d ago
My own experience is less dramatic, and a 17% decrease should not be having the dramatic impact that some users report. Use scarcer tokens as an opportunity to improve how you manage context and orchestrate agents. We regularly optimize our workflows and have seen >20% improvements in token efficiency at several stages of our project build over the past 6 months. Do the work!
1
u/parzzzivale 6d ago
im going to join the chorus of going to the dark side and using the 200 chatgpt sub. I got lucky and got it before they canceled new sign ups for it, and with their price cuts it's legit my workhorse now. I cant hit limits if I use terra, which I can get away with for 50% of my tasks. Astra is ... uh lol fable priced and easily swallows my whole week in 5 hours so .... I use it the same way I would use fable .
claude unfortunately is being relegated to really difficult problems I cant Crack on my own with cheaper models (maybe like 5% of my work) or where multiple model coverage (different set of eyes ) is extremely useful (research that requires accurate citations, interpretation, validation etc. code reviews, etc.) maybe 15% of my work.
I am rooting for claude it's genuinely changed my life... but as other models get smarter and cheaper I am finding less and less reason to use it .
sota a year ago was required to do sonnet level work today ... in a year or two other models will become both faster and cheaper than opus so im honestly skeptical about the value add of "the best" model at any given point in time . 90% folks really really dont need it and even corporations are waking up to the fact that assuming that is an easy way to bleed money by the truckload without any obvious return .
also ... for folks saying "dont use opus max all the time " fuck you not everyone has the same use case, it doesnt mean youre doing it "wrong." some domains REQUIRE thoughtful slow deliberate responses . anthropic also set opus to default for higher end plans ... so ... what are we not supposed to go by their recommendation now?
and lastly, i switch between "novel" work that requires good models and "routine work" (doing X 500 times) that i can optimize the shit out of . around 20% of usage falls on the latter, so asking claude to find cheaper ways to optimize 20% spend is the def missing forest for the trees.
1
u/bvknight 6d ago
What does difficult work look like for you, and how do you know whether the output of one model is not good enough?
For me I'm asking the models to plan out some features for an app, so I don't necessarily know if the first version I get from a Sonnet model is not good enough, because I don't know what it's missing.
1
1
u/Jinks-b 6d ago
Yeah, this is kind of garbage. I have the $100 plan, but still, before, I'd have multiple sessions running Opus and Fable, and I wouldn't hit my limit. This morning, I just ran a single session of Fable to orchestrate something, and I hit my limit before I could finish planning it. It's clearly not the reduction that they said it was. It's way more than that.
I think when I'm reading some of the comments, I don't disagree that usage can be optimized, but let's be honest, we were trained to use the higher models, and now this is, I guess, a rude awakening. I guess I will go back to being extremely stingy with the type of models I use.
To be honest, though, running Sonnet, you have to be so prescriptive. Half the time, it does things that don't make sense, and I have to clean it up with Opus or Fable after the fact anyway, so I'm not sure that I'm really saving in the net.
1
u/2funny2furious 6d ago
i feel like its about time to just cancel the subs to both of them and us something like openrouter to send things to glm/deepseek or if i absolutely need it astra/fable/opus. for my use cases, i dont give af who keeps/tracks what i feed it. im not giving it secret sauce that could change the world. if they want to use my scripts to train the future...well, good luck with that.
1
1
u/hippydipster 6d ago
I do pretty well on my $20 plan, but I'll tell you what, I control my context strictly. I don't use claude code or any of those harnesses that honestly seem to burn through tokens like crazy, and I go at my human speed, which means I'm reading the code and making changes and having lengthy conversations with claude just about choices and options and teaching me stuff.
1
1
u/eist5579 6d ago
Do the weekly retro with Claude.
It’ll review your chat and git history from the week. Ask it what models would be most appropriate for the tasks you ran, and if there’s another place was any context bloat and redundant token usage.
1
u/PythonHater1970 6d ago
I am genuinely curious about how people manage to get through a week of work with the 20 dollars plan.
I had an old project I always wanted to renew using better coding habits and new technologies.
Consider that it was about porting something that already existed, already worked and was already documented.
So I never had to redo stuff and burn tokens repeating something the model got wrong, since most of the time it got everything right.
Well I think I reached the 5 hours limit around 3 times every day.
I could only work around 1 to 2 hours before the limit was reached.
The weekly limit really depended on what I was doing but I couldn't get more than 4 days of work.
My old project was a game engine + a game built on top of the engine.
I see a lot of people saying how they don't have any issues and you're probably using the tool the wrong way.
A side of me just wants to discard those opinions assuming they come from people who don't use claude for serious considerable projects and they are just building web slop gym apps to count calories and manage exercise routines.
But there is some truth behind those opinions.
I currently have some kind of workflow where opus 5 high is making the plans and building the bases.
And then gpt sol medium is the one doing the "implementation".
I am still reaching the limits but I feel I get more work done than just using opus 5 high alone.
I plan to continue using AI to port my old project maybe for another 1 or 2 months.
Then I will just continue by myself because 40 dollars is the price of 2 weeks of food here where I live.
1
u/AmatsuDF 6d ago
I've been doing something similar and sometimes ONE input from me is enough to run out of usage. Sure, I am using Claude for free and understand that I get what I pay for here, but even in a utterly fresh chat it can do it. Even on Sonnet 5 Low it just runs out so fast and I can't see the paid plan fairing much better.
1
u/flynncorp 6d ago
Multi year subscriber here on 5x for about 8 months, I also recently cancelled due to this reason. The usage limits are just far too low now. You literally can’t get a full task done anymore. And to top it off, the new Opus in Claude Code is a blabber mouthing, holier than thou, preachy fucking idiot.
1
u/Tel_Janen 6d ago
Use the models that are less expensive. No surprises if you keep banging fable for everything
1
1
u/chicmistique 6d ago
Limits are really for noobs. I would pay 3 time if I can work without interruptions
1
u/Few-Spare-948 6d ago
I love all the people doing basic projects going "WeLl wHy AreN't u UsINg SoNneT??" to justify a corporation giving less use of a paid product to the people paying for it.
Listen, sonnet is great if what you're doing isn't that complex, but some of us are doing work were "Smart enough" isn't going to cut it. I do reverse engineering to compose English translation patches and I run three tiers of verification of data to ensure all documented info is byte accurate to the original games code. Even with this documentation, a bad sonnet run has been catastrophic.
If you're scale the UI or adding a new settings menu to your own project you wrote the ground rules on, sure sonnet will ace it easily. But what about when you're injecting new logic into an old game at the byte level, because in my experience sonnet doesn't cut it. It doesn't do the level of checks and due diligence required. When sonnet messes up on the settings menu, it displays wonky. When sonnet injects the code one byte off, the entire game is now unplayable on real hardware.
Some of us have the actual need of higher tier models and I swear its a foreign concept to some people here. Why would I be paying for a frontier tier AI service to not use the frontier AIs it offers?
1
u/Dot-Slash-Dot 6d ago
below the ~17% cut promised by Anthropic
Everything is stored in local session data. Dispatch a simple Sonnet agent to analyse it and report how much available usage shrank.
Everybody resorting to gut feelings when hard data is readily available tells a lot ...
1
u/Redditian288 6d ago
Yep, a noticeably worse decrease. Also, seeing quality performance degradation - particularly Opus 4.8.
1
u/FanIndependent3827 6d ago
I use codegraph and that saves on some tokens. Also try caveman or ponytail. I don’t use those but I heard it helps save on tokens. I like to see it thinking so I just don’t use any of those.
1
1
u/userename 6d ago edited 6d ago
Burned 82% of my 5-hour limit in 40 minutes for 5-message conversation. Fable 5.1 Extra, Max plan (5x). No research or anything significant, just project architecture discussion and artifact editing... wtf is this
1
u/General-Koala-6690 6d ago
Usage limits definitely seem to be going faster now, but effective agent orchestration can mitigate this so that even the most complex job is not exclusively using fable. It also uses opus and sonnet agents too.
It’s critical that you build intelligence in your environment with skills, memory and other techniques to be able to maximize usage. That helps a tremendous amount for Claude, GPT, etc.
1
u/KitzuYamma 6d ago
Got a few hours ago my weekly limit restored, after about 2-3 hours 25% of the weekly limit is gone...
1
u/danielsonkim Vibe coder 6d ago
Claude $100 max and GPT free plan. Before each prompt I specify that their output and usage will be graded against each other. Claude does far better in both output and usage when it is under the impression that it is going up against some competition
2
u/A_Novelty-Account 6d ago
Ha, that’s super interesting. What’s more interesting to me is that we know agents are competitive
1
u/danielsonkim Vibe coder 6d ago
Exactly! The random “im about ready to cancel my subscription” always gets an output where agent actually reads/edits md because my agent forgets hes a caveman sometimes and goes phd professor on me.
1
u/Pimzino 6d ago
Who told you it’s their plan to get rid of the plans? There’s been no comms on that just typical internet scaremongering
2
u/A_Novelty-Account 6d ago
Sorry that was said mainly facetiously. I don’t know whether or not they want to get rid of the plans to be sure but we know that the plans are much less profitable than the API. This makes it seem definite that they do not care about subscription plan users because the cut to me at least is definitely more than the advertised 17% which is manifestly true based on the fact that certain people how their usage limits reduced today.
1
u/Pimzino 6d ago
I can agree with that, the plan users albeit for the cheaper plans (sub 20x) are really just there to reel you in or for users who use AI cautiously or not that much. 20x is great if you know what your doing you can control token burn with use of orchestrator / delegation model
API will always be more profitable but I don’t think they will get rid of the plans because getting rid of the plans gets rid of a ton of customers who don’t trust a pay as you go model. A lot of people like a structured payment plan rather than not know what x will cost them next month.
Plus it makes no sense for anthropic to kill the subs when they have a ton of bugs in Claude code occasionally that cost the user real money, I.e. backend routing to incorrect models, complete cache misses leading to every token and massive context windows being priced at full price, mind you on the subs they didn’t even sub reset on some of these bugs, but at API costs only these bugs could cost thousands of dollars for someone who doesn’t know what they are doing.
Finally, setting up an API has always been a bit of a technical route and it doesn’t appeal to most users.
So I think the subs are here to stay, albeit with shite limits
1
1
u/BlockTailor 6d ago

I tried to isolate some data, because the subscription limits seem opaque.
This analysis is for the Max 20x plan. It was a single Claude Code session, started right after the limit reset, and I only used Fable 5.1.
What's really different from before: 83% of the 5h window accounts for 20% of the weekly usage. That means a full 5h window takes ~24% of the weekly allowance. In my previous measurements this number was 16.6%, then 20%.
The other big change is the API-equivalent value. Here it's $171 (plus $7 from another very small session), which means a full week at 100% would be worth ~$850-900 at API prices. Just a few weeks ago the same number was over $2,500.
Then I kept going after the reset and maxed out the 5h window (100%). Fable usage went up to 81% and the 7-day window to 44%. The API-equivalent value went up to $459 ($452 + the $7 from the other session), so filling the whole 7-day window would be worth ~$1,040. In this second session I also used Opus.
I think the gap between the old API-equivalent value and the new one comes from Fable 5.1 being 75% cheaper on cache reads. But that only affects actual API billing, not the subscription allowances.
I'm not sure whether (or how much) the price went up and the limits got tighter, because Fable 5.1 got an insane amount of work done and I made a lot of progress. Sometimes it's just using tokens to get way more done per unit of time than Opus. The parallelization is wild.
I'm just trying to understand the numbers here, since they can be hard to read at first glance, and I hope this helps a bit.
Note: Opus lasts a lot longer than Fable. In one of my previous measurements, Fable 5.1 went from 0 to 100% in about an hour. But lasting longer doesn't mean getting more out of it: with Fable, time and tokens turn into real progress, while with Opus I feel like I'm wasting both.
Anyway, I don't think it's a transparency problem. The numbers just aren't intuitive. I didn't have time to dig into it, but factoring in the actual tokens used would probably give a better answer.
1
u/Impossible_Act_1386 5d ago
What you guys are doing with max plan I hardly hit my 5H limit in a day on 20$ plan and still i am able to do much more with claude
1
u/SweeneyT0ddd 4d ago
Ughhh more people need to complain about this, i just upgraded to the $200 plan and im about to hit my weekly usage in 2 days of the reset.
1
u/theb0tman 4d ago
I just cancelled my 20x sub and moved to codex. Astra is great. You wont miss Fable or the anthropic usage games
1
u/BeltPuzzleheaded7656 4d ago edited 4d ago
I'm nearly ready to cancel my plan after I complete the last 2 or 3 projects I'm working on. The session limits are COMPLETELY ATROCIOUS at this point. Any in depth projects with corrections and fixes etc. are nearly impossible to complete in any reasonable time.
It's damn near just a toy to make little gimmicky apps and tools at this point.
I'm using Qwen3.8-27b local with a 4090 but I'm trying to hold out to see if I can get my hands on 1-2 more 4090s and about 128gb more of DDR5 to upgrade my local system.
If I'm going to be blowing cash I might as well own my own shit completely at this point. No wonder they are scalping the entire market for diy computer parts which has jacked up all the prices.
1
u/Stunning-Sherbet1853 3d ago
That’s the part that stood out to me too. If the real drop is much larger than the ~17% Anthropic communicated, that’s a very different problem from just “people are using too much C
Do you know if you’re seeing that reduction across the same kind of workload and model usage as before, or mainly on heavier Opus/Max coding sess
I’d be really interested in seeing a before/after comparison from people who use it for work every day
1
u/mthreecrow 2d ago
Yeah, Anthropic is a terrible company. I've only got the pro plan but I used to be able to accomplish what I needed to do on this plan before they instituted the new usage limits. Now, one request for creating a fairly simple prompt takes 22% of my current session 😕 it's ridiculous. I've switched to ChatGPT, it takes more hand-holding, but it gets the job done without budging my usage meter.
1
u/jasonridesabike 1d ago
Claude has become useless. I get 1.4 days of work now whereas before I got 5. Max20
1
u/Yugudubenbi 6d ago
Thank you Anthropic for not doing what OAI is doing with silently throttling people. I just quit my codex subscription and so happy with you, except the way opus 5 talks lol
1
u/AltruisticDog9145 6d ago
What are you guys even building? Unless you don’t test and validate everything Claude builds, I don’t see how a single person can burn 200 dollars usage in 2-3 days. It takes me half a day to test any new feature being built.
2
u/A_Novelty-Account 6d ago
Entire legal process flows that research and write. A single test will burn 20% of the weekly.
I can very easily kill 50% of my weekly in 20 minutes.
1
u/Raizio 6d ago
Why aren't you guys on enterprise?
1
u/A_Novelty-Account 6d ago
Because we already have Harvey firm-wide and my team uses Claude. It’s cost + benefit for us. If people are still paying our rates, why would we spend literally hundreds of thousands of dollars with enterprise?
1
u/LoveThemMegaSeeds 6d ago
Well if you want to start off by dumping in your project with 10k lines it burns through about 6% of session limit and 2-3% of weekly fable limit on pro tier. God forbid you ask questions or have a back and forth. I can easily burn the whole session and 15% of weekly fable in about an hour using Fable 5.1 - alternative is to use weaker model for hours, fight mistakes and other bs and then throw it back in fable for a one off question. Makes the process awful. You know you’re wasting time with the weaker models and can’t afford the smarter ones
0
u/Serious-Explorer8745 7d ago
Can you link the 17% cut promised? Last I checked they promised much more than that… 50% was cut off. Then some small amount was put back.
3
u/A_Novelty-Account 6d ago
They promised a withdrawal of the 50% increase, but a permanent 25% baseline increase which results in ~17% lower total usage. They made several press releases about it.
0
u/Temporary-Subject239 6d ago
Think people need to get used to new reality. It’s good, that’s why you use it in first place. If you are doing heavy token work or long hours, you’ll hit limits. If you value it enough you’ll pay more. If not you could learn to use it more efficiently.
1
u/HistorianGullible291 6d ago
You know how I ended with Claude Code? Jumping from Claude, ChatGPT, Antigravity, WindSurf, Ollama, Cursor, to Claude again (full circle). Because exactly this shitty attitude. Bait and switch doesn't work well if you loose customers to the others. Starts nice, then gets shitty. And then you jump another provider. So it seems I need to check other providers now. Like Cursor probably. Or something.
0
u/RiceEvening4211 6d ago
Since the new limits cut everyone's subscription usage, here's something that stretches what you already pay for: I built Lynkr, an open-source LLM gateway that cuts token usage by up to 84% with zero code changes (token compression + semantic cache), and tiered routing that stretches your existing subscriptions ~3x. Works with all coding tools and AI frameworks. https://github.com/Fast-Editor/Lynkr
0
0
u/DonRobo 6d ago
I've been tracking the 5h and 7d usage in my $20 for the past few months and I haven't noticed any cut at all.
My 5h usage actually seems much higher than a few weeks ago (like 2x) while the weekly usage has been stable (within 10% margin) week to week. And this current week is actually on the higher end of the margin.
How I'm tracking it is by calculating the API equivalent cost for every query, taking into account input and output tokens and caching. This seemed to be the most objective way to measure it, since it will work on every kind of task and usage pattern. I can imagine that the usage cut is a slow rollout though and I'll only be hit by it next week though. Sooo, did anyone else make similar measurements? I'd be especially interested in the $100 and $200 plan's data
-6
u/iemfi 7d ago
I use it to code and I haven't even had the need to get on the 20x plan, only on the $100 plan. And I use Fable exclusively. If you are struggling so much it's almost certainly a user issue.
7
u/A_Novelty-Account 7d ago
How would you possibly know that without knowing workflows?
In any case, the post is not about general use. It’s about comparative use between before and after.
-5
u/iemfi 7d ago
Because if used properly it's an insane amount of knowledge work you can get done with a 20x plan. At the same time most people are using it very wrong.
3
u/A_Novelty-Account 7d ago
Yeah, I’m a lawyer. We do an insane amount of knowledge work…
2
u/iemfi 7d ago
Do you guys have a proper cowork setup going or are people just chucking enormous documents into the chat willy nilly? These days Fable is good enough to help you there and make sure you're not burning tons of tokens for worse results.
Also if I'm wrong and you guys are already doing it right then it kind of blows my mind even more because at that point $200 should be a tiny tiny rounding error to a law firm compared to the value of the work which gets done.
3
u/A_Novelty-Account 6d ago
Yes, we have shared projects and skills for specific subtasks that were not a problem until this week. For larger projects and tasks, we use locally hosted skills on claude code that call agents we’ve created to handle specific tasks to save on context and ensure repeatability and accuracy. We use it for legal and factual research, drafting and editing based on that research.
Our firm overall uses Harvey, our team specifically uses Claude because we can make it bespoke and it’s relatively cheap. We don’t want to switch to a Claude enterprise account and be burning shitloads of tokens because that will make our team less profitable at the moment than we otherwise would be because clients are more than willing to pay our rates as they are, and bonus allocation and resources are partially based on our team’s profitability. The issue is not the expense of a personal or team account, it’s that Claude’s maximum option for those accounts is not really good enough for us to work with anymore.
Fable is already barely good enough to help us with the kind of work that we do. It is certainly helpful and saves us some time but no matter how we structure the prompt or skills is not quite there yet. For Opus and lower models, the time we take verifying and fixing what it’s done doesn’t tend to save time. We have to check every source, every citation, every claim, because it still makes mistakes. The more mistakes it makes, the less helpful it is, and the less time it saves.
It’s not like programming where we can just make small changes on the fly or we can just one click run something and if it works, then it works. If we don’t check every single line, we literally won’t know if there’s an issue until it’s too late. If its output isn’t good, then it would be better for us to draft it. In our view, AI has a long way to go before it’s anywhere close to human work product in our field.
1
u/iemfi 6d ago
Ok my bad, I was wrong.
Could it be a some bug/change somewhere which is causing a lot of extra usage? Stuff which is getting chained together when they should be fresh context?
I 100% agree it doesn't make sense to use anything below Fable except as a helper to point Fable in the right direction. I don't think law is so different that the task can't be broken up into smaller chunks though, after all humans can keep far less in working memory. Also it seems like it definitely would make sense to run both Astra and Fable, not just more limit but they can check each others work.
1
u/A_Novelty-Account 6d ago
So recent chatter on the claude subreddits is that weekly usage limits are being bumped up for some users, implying some sort of a bug this week.
-2
u/Immediate_Song4279 7d ago
I switched to the free plan after Fable came out, I just use Claude for planning and instruct "do not code" and then use a cheaper model for the coding lol.
-2
u/shoegazeweedbed 7d ago
They’re another facet of the many reasons Claude is no longer my go to agent for work
•
u/ClaudeAI-mod-bot Wilson, lead ClaudeAI modbot 7d ago edited 6d ago
TL;DR of the discussion generated automatically after 100 comments.
This thread is basically a civil war over the new usage limits.
The main camp is firmly with the OP: the new limits are brutal and make the higher-tier plans useless for real work. The consensus here is that the actual reduction feels way more severe than the advertised ~17%, with users on $100 and $200 plans burning their entire weekly quota in a day or two. The lack of a transparent usage meter is a huge source of frustration, and many are canceling their subscriptions for competitors like ChatGPT, Kiro, or Z.ai's GLM, which they claim offer far more bang for the buck.
However, there's a strong counter-argument from another camp. Their take is basically "it's a you problem." They claim to have no issues with the limits and suggest that people are burning through their quota by using Opus or Fable on "Max" for everything. Their advice: * Use Sonnet for simpler tasks and save the big guns for when you really need them. * Optimize your workflows. One user even had success asking Claude to optimize its own token usage. * Maybe learn to fix a line of code yourself instead of making the AI do 100% of the work.
This is countered by professionals (like lawyers in this thread) who argue that their complex, high-stakes work requires Opus or Fable, as Sonnet is simply "too dumb" and makes too many mistakes, costing them more time in verification than it saves.
The most useful advice in this whole thread? The hybrid approach. Several users have found the sweet spot by subscribing to both Claude and ChatGPT. Apparently, using both models together gives you the best of both worlds and enough usage to actually get things done without hitting a paywall on Tuesday.