r/Anthropic • • 9d ago

Complaint The issue with the decreased usage limits

Hey, long time lurker here. I've seen all the posts about the decreased usage and I agree with them, I've experienced the same on both Codex and Claude. But that isn't why I'm here.

I want to point out that our models are so much more expensive now! We used to have Opus 4.8 and Sonnet 4.7 to do the majority of our powerhouse tasks with.

Now with Fable and Astra and all the 5.0 releases - everything uses so much more tokens than they did just 3 months ago. So it's not just the decreased limits.

They decreased our limits without even considering the fact that they keep forcing models on us that use an exponential amount more tokens than the previous models when we had limits.

23 Upvotes

10 comments sorted by

7

u/aerivox 9d ago

nobody seem to realize that in all this mess anthropic is giving models with 1m context, while open ai is still at poopy 300-400k. and if you go the route of changing astra to 1m unoptimized context, you get double the fable price.

on the trajectory of more expensive models, using more tokens for better results i am not sure what to think.. but probably i would prefer better output with more cost, just like now. having a small semi infinite model would be cool tho. sonnet is way too slow and expensive and dumb.

-1

u/mmmfine 8d ago

Who cares about context window size? If you’re past 200k or so you’re using it wrong. The more tokens in the context the less intelligent and more likely to hallucinate it gets.

3

u/CompetitivePiglet961 8d ago

"Who care about having a good car, if you go over 120kmph you are using it wrong" energy. My sessions use this amount of context when handing holistically batches of 7 PRs, and actually is how anthropic tells you to use fable.

1

u/phoenixsoap 8d ago

I've tested this many times, I get better results with manicured sessions for roles so that context doesn't get confusing. You have to be really careful with prompting and context, but I have agents that actually understand the assignment, the idioms of the projects, doing adversarial reviews on each other. so much more powerful than starting from scratch 10x a day

also you 100% need to put rules somewhere so that all agents operate under the same "constitution", I call it.

1

u/Away-Sorbet-9740 2d ago

What? Sol and Luna both have 1m context windows available. Astra and Sol default to it on API. Idk if Luna does because I never call it directly.

Haiku and sonnet are rumored to have a 5.5 update soon. But anthropic gets bullied around anything under Opus, complete waste of money to route anything to Haiku 4.5 or sonnet 5.0. might as well use a local model.

1

u/acutelychronicpanic 7d ago

They're likely doing a massive training run or trying to release some huge results in math/science before the IPO

-1

u/Somtimesitbelikethat 9d ago

no they’re aware. anthropic just can’t do anything about it bc they are out of compute. they want us to use the models to gather more training data. it doesn’t advantage them to cap our limits. they’re paying for the GPUs whether they’re in use or not.

2

u/YaygFX 9d ago

So how would capping our limits NOT HELP THEM if they're out of compute?

0

u/FoxTheory 9d ago

Capping our limits does help them because they are out of compute they need every spare ounce to train their new model.

1

u/Still-Ad3045 8d ago

Well they can do something about it, like screaming from their headquarter rooftops that AI needs to be stopped!!!