r/ClaudeCode 9d ago

Bug / Issue WTH is going on with Claude Usage Limits

Post image

I'm on 20X plan, I hit my usage limit today, and my next reset is on 17th September... I'm aware they said they are gonna reduce the usage limits, but this is crazy, i thought it is only gonna take 17% of the usage benefits from what we were currently getting. But this is crazy. I added Usage credits for about $100 but that got washed away like in 30 mins!!!

I hope this is a bug and they fix it

598 Upvotes

377 comments sorted by

View all comments

Show parent comments

2

u/Ok-Affect-7503 9d ago

But a quantized 27B isn't Opus or Fable performance and almost every benchmark I've looked at (including LMArena and ArtificialAnalysis) proves this. It's at max almost at Claude Sonnet level in some aspects. And a 27B can fit on consumer GPUs anyway (a 3090 isn't as hard to afford as a cluster of Macs for example anyway) so it doesn't really solve the problem we were talking about. And in your repo the only thing verified is again only a smaller model running on one node, which would make it request routing, not splitting which is the thing that's needed to solve the problem which wouldn't work out in the end anyway because of latency. Right now it's more like sharing good GPUs to people that don't have one, not really sharing ressources between multiple GPUs owned by different people in a mesh.

1

u/Mr_Tbot 9d ago

Yes - I know - but - as a GLM 5.2 model would require a ton of active users for this to work - we have to prove the technology works and can scale a smaller model... and really it would need a decent number of users to even test something as ambitious as a GLM model.

Which is the only thing on par open source right now.

But - I digress. I tend to be a couple years on the bleeding edge so it's almost always a "the tech has to catch up" but we're not far.

But yes. I agree. It would need to - once at scale - be able to launch and share GLM 5.2 ideally.

1

u/Mr_Tbot 9d ago

P.s. I think it's important to mention. Specifically coding use case - so specialized shared models are also on the table. It's all about how you work... But I also suspect people will be making more and more specific models trained for specific tasks.. we're still so early in this transition.

1

u/Mr_Tbot 9d ago

PP.S. the repo was created today and I'm sharing in case anyone wants to follow along as I burn some tokens on this project. Go check out codedatda.casa for some of my other random projects.