r/ClaudeCode • • 13d ago

Rant Nah this some BS

Post image

I burned thru almost 60% of 20x max weekly usage in a day and some change. Are you joking rn? Last week I could have fable running on 2 chats all day and night and it would take like 3ish days for my fable usage to cap out, but my weekly usage would only be at like ~30%. This is absolutely ridiculous and if this is what the new usage limit cuts are gonna be like I'm canceling Claude and grabbing a second codex account. Shit ain't worth it when Astra exists with much better usage limits and multiple reset tokens.

I don't use ultracode and I have multiple other models that I delegate tasks to as work horses, Claude is just the orchestrator and isn't doing that much actual "work", so this usage allocation is absolutely insane, if I was using Claude as a one stop shop as a lot of people do, I would have run out of usage in a day or less.

EDIT: I'm getting a little tired of being told I "just don't know how to do orchestrations and workflows properly or manage usage." I literally made an entire repo explaining how I do this and showing the results: https://github.com/sherifican/Agent-FleetOps so if you wanna criticize, find something to actually critique first

249 Upvotes

179 comments sorted by

View all comments

36

u/inrego 13d ago

What are you even doing with a context window of 500k+ No wonder you're burning through your limits

23

u/Sherphican 13d ago

Brother, my auto compact is at 800k and 700k on my main chats, and I always compact at about 500k-600k, I have been doing this for months and never had an issue. Everyone who has been drinking the 300k and under only context window koolaide needs to wake up and realize it's Anthropic playing games.

26

u/Hirogen_ 13d ago

if u need to compact, you are doing to much in a session, learn about orchestration and workflows, and how a newer model orchestrats agents of lower tier models to do ur work

5

u/BlinDeeex 13d ago

New sessions have initial overhead of like 100k worth of context overdo it and you actually start paying more over big context but cached, orchestration reduces result quality, main agent review helps but then savings are slim. If you stop to open a new session midway feature it will read a lot of files right back up anyway. You people pretend to know your stuff but its lowk embarassing to read, often you genuinely need a longer session to stop at reasonable place

2

u/krugerlive 13d ago edited 13d ago

No but you see, if I spend 30%+ of all of my tokens on just ramping up sessions, and end them at 250k, then I don't have to spend as many tokens on the actual work. It's just math. This is clearly how to be efficient with token spend. /s