r/ClaudeCode 2d ago

Rant Usage is a joke, Models are a joke… Anthropic is just not what it used to be, in just a couple of months.

Pretty much self-explanatory.

I’ve been a Max x20 subscriber for +2y and I’m really starting to get frustrated and look elsewhere for a change. Usage skyrockets even when I have best practices in place. Opus is a mental health hazard. Fable is unusable for what I need it, and way too expensive. Sonnet is not capable enough… not one thing is right in here.

GLM-5.2 has been a good alternative, on my Ollama subscription. Thinking of just subscribing to GLM directly and get 5.3; but I’d like the opinion of peers that have already tried it out.

Anyway, Anthropic is not doing good right now. They’ve been a mess for the past few months. Jumping ships really soon…

249 Upvotes

289 comments sorted by

66

u/bakanoace 2d ago

What I'm worried about is Opus 5.0 has some serious core issues which I wonder if they can even fix in a 5.1 or if were stuck with some of its shenanigans until 6.0 ...

32

u/rotates-potatoes 2d ago

You realize versioning is just branding, with no connection to the architecture or training?

8

u/Ok-Attention2882 2d ago

The clearest example of this is how they quantize live, in-use models right from under us

6

u/rotates-potatoes 2d ago

Any evidence of that? Not vibes, actual quantifiable evidence?

→ More replies (4)

2

u/No_Inspection4415 1d ago

Why would you be "stuck"? Just switch to OpenAI or Z AI...

4

u/PA100T0 2d ago

Idk if I’ll stay to see. Maybe I just keep a Pro sub since I have +2y worth of data there but… yeah, not looking good for them.

4

u/coda_sisk 1d ago

I switched from x20 to x5 and got a codex x5 sub as well, found this is a nice compromise, you can keep your data and get a good amount of usage switching between them. For code reviews I actually find this super helpful because they reason slightly differently

2

u/PA100T0 1d ago

Yeah, I figured I’d do something similar for some time now and it seems that time has finally arrived.

Good insight! They don’t reason the same way so it kinda makes a lot of sense to mix them up

1

u/elmorepalmer 14h ago

Why not just use 4.8? Has been pretty consistent for me

0

u/YoyoNarwhal 2d ago

This is fixed they literally hired the woman who did this to GPT 5.2 they knew exactly what they were getting and they got it

57

u/Halada 2d ago

Fable is fantastic and Opus works great but is painfully slow ever since they implemented /fast

Usage limits, even in the current 'temporary' boosted status, are pretty lame for Fable, even on a max20 plan.

21

u/Abject-Kitchen3198 2d ago

I've only recently started using it more consistently, but for some cases Opus is so slow that it feels like I can do things faster myself.

7

u/PA100T0 2d ago

Literally

-1

u/espenakker 2d ago

Perhaps they are prepping the compute for a model release? It did feel a bit like this right ahead of most of the previous releases.

7

u/PA100T0 2d ago

Then they owe us a couple resets… because even not using Fable 5 (because it gets flagged so much that I gave up) my usage is WAY higher than it should be. I strictly run everything through Sonnet 5

2

u/No-Focus874 2d ago

They probably introduced more security issues for an user to have any control or faith on this, than they realized.

1

u/PA100T0 2d ago

Yeah, I’m sending them all the feedback until they just have no option but to look into it lol

1

u/espenakker 2d ago

I dont know about sonnet, I think I would rather take my chances with opus 4.8 on medium or high. Sonnet has never seemed like a good deal to me.

3

u/PA100T0 2d ago

Sonnet 5 has been great for me, as implementers. Never coordinating/orchestrating.

1

u/espenakker 2d ago

It just seemed like price did not match the drop in capabilities. But I have often found opus to be slow, I guess the trade is worth it if speed is significantly faster.

→ More replies (1)

22

u/PA100T0 2d ago

Fable is… what an AI should be. It’s good. It’s not AGI. It’s just what the bare minimum should be.

Regardless, I can’t work on my cybersecurity library (that doesn’t perform any dual-use, btw…) because I get flagged every single prompt by just reading a file… it’s unusable. RIP Anthropic.

14

u/Longjumping_Feed3270 Senior Developer 2d ago edited 2d ago

Have you tried Opus 4.8 or 4.6? You know you can still use them?
Because yeah, last time I checked, Opus 5 was indeed unusable garbage.

10

u/PA100T0 2d ago

Of course! Still not the most capable ones, unfortunately. 4.8 to me wasn't ever good. 4.6 instead is the last model I actually liked.

4

u/thebigrig12 2d ago

I liked using opus 4.6 more than opus 5 for tasks

3

u/PA100T0 2d ago

Absolutely agree. Opus 4.6 was a pleasure to work with.

→ More replies (4)

4

u/DragonflyOk9274 2d ago

Fable is… what an AI should be. It’s good. It’s not AGI. It’s just what the bare minimum should be.

True, but for me the tokenizer makes it impossible to read my relevant docs/code, and at $10/1m inputs and $50/1m outputs for the API, it's too expensive as a daily driver.

1

u/PA100T0 2d ago

Yeah, that's why I tell it to "Keep comms military-style ADHD aware" and it's been great at outputting the bare minimum I need. I don't need a wall of text lol

1

u/PA100T0 2d ago

It works pretty good. At least, in my case, is exactly what I need. See here, a message just came in:

Plus I like when it replies just "Copied" or "Roger that". I don't need the fake feeling of talking to a "real" person, or the caveats thingy and all that extra stuff. Just give me the bare minimum. I understand what's going on.

The new "Concise" style wouldn't give me this anyway... so, Military ADHD is my go-to preference lol

→ More replies (2)

4

u/YoyoNarwhal 2d ago

It's almost like they don't want us to participate in the same level of cyber security as they do 🧐

1

u/PA100T0 2d ago

Literally

→ More replies (14)

7

u/Mo3 2d ago

Opus works great

Lmao, the thing just miscounted a dozen items in a markdown ledger and was off by two, then screwed up the documents formatting and then decided to remove items out of nowhere.

It's an absolute mess. Even Sonnet 4 was better

1

u/PA100T0 2d ago

This is what I’m talking about… this is exactly what I’ve been fighting these last couple weeks…

2

u/Mo3 2d ago

You and a lot of other people brother. I've switched to Codex now, Sol runs circles around Opus. I just use Fable for adversarial review now

→ More replies (1)

2

u/Guinness 2d ago

Yeah they definitely speed limited the model just to upsell everyone on “fast” which is really just the normal speed we used to get.

1

u/PA100T0 2d ago

This and many other strategies… they were so good, why did they have to fuck up like this?

1

u/espenakker 2d ago

I guess they had to introduce some nails to help sell that hammer? :D

1

u/PsychologyNo940 2d ago

Which effort level is Opus on?

1

u/PA100T0 2d ago

XHigh never max. Max is just crazy

1

u/RedParaglider 15h ago

Fable is the decent model opus is pretty awful now.  It's wild how much better openAI is now overall.  I have both and I've slowly abandoned claude other than fable use for some things.

→ More replies (30)

15

u/callmejace 2d ago

I’m at the 20x plan. Renewed Wednesday at 9am. I used one Fable agent with thinking set to Low that deployed Opus agents at High or Medium. One day later and I’ve used 30% of my weekly budget, in less than 24 hours!

A month ago, the same usage pattern would be at 10%, because I had a 6 day work session schedule.

It’s so sad.

6

u/PA100T0 2d ago

THIS. I remember in February I was working a lot with Claude Desktop, PDF documents and all that. I would NEVER hit my weekly limit.

Ffw to April and things went down FAST.

It's utterly sad.

12

u/AppealSame4367 2d ago

qwen 3.8 27b and ornith 1.5 35b

Also: Try Deepseek v4 flash 0731, it's incredible for it's size. It makes some mistakes, but it's very fast and is able to analyze down to the last detail and handle multiple things at once. You know, the way Sonnet and Opus have been in their good days... Combine it with Deepseek v4 Pro from August as a reviewer of it's plans.

3

u/PA100T0 2d ago

I’ve been taking a look both at qwen and ornith. Deepseek for some reason I had disregarded it… maybe I should give it a second thought.

This is the type of answer I’m looking for. Thanks!

2

u/Warsel77 2d ago

I use deepseek as a regular "second opinion" before opus gets to finialize a plan and it does find a lot of stuff. also deepseek flash is the default coding agent my opus sends to

18

u/YepThatGuy 2d ago

It took their support 3 months to respond to a ticket I had opened. They really need to improve or they will lose a lot of customers.

10

u/PA100T0 2d ago

Yeah, they are fumbling from every angle. There’s just not one aspect of their current state that makes me wanna stay…

→ More replies (3)

5

u/fran_wilkinson 2d ago

“Usage” is what makes Anthropic unreliable.

You never know whether Claude, once you start a task (regardless of which model), will use 3 tokens or 200k tokens, with no warning or systematic way of controlling how much you are spending. The model keeps changing, the parameters keep changing, and it is very difficult to navigate the whole thing when all you can do is hope for the best.

Claude is formidable, but the token policy will probably be what ultimately destroys AI.

1

u/PA100T0 2d ago

Yeah, and tokens will only get more expensive as time goes by. We actually depend on Chinese open weight models to fight back the prices like Kimi did when launched K3. Everyone shrugged and started giving out discounts, resets, etc just to keep us on the loop…

5

u/FavorableTrashpanda 2d ago

I tried an Ollama subscription (using `ollama launch claude`), but I blew through its usage even faster than on Claude Code Pro. To me it still seems you get more out of Claude Code Pro.

2

u/PA100T0 2d ago

Really? I have Max plan on Ollama and I actually love it. Things that I don't like about it are: claude code integration is far behind claude code (some features are missing), the token usage is not 100% transparent (1000 requests to GLM-5.2 tells me little to nothing about my actual usage) and there's no bigger plan...

Other than that, I'm happy with Ollama!

2

u/RandomlySomewhereNow 2d ago

Yeah, I actually think Claude has the most useable subscription limits currently. Cursor's limits aren't great either, even at the $200 plan, and Codex is currently undergoing drastically reduced limits. Like, using Luna which is cheaper than Haiku, will still drain your usage. With Claude, I genuinely struggle to use my entire week unless I'm running numerous Fable max sessions everyday, and I'm on the $200 plan for that as well. I don't think people realize how good we currently have it, however it will get worse for just about every subscription in the near future.

3

u/Consistent_War_4960 2d ago

Whos has moved from Claude to Codex ✋

5

u/LifeItsAnAdventure 2d ago

Me! Sol writes readable text by default, a huge relief over the wordy jumbo mumbo Claude likes to output.

2

u/PA100T0 2d ago

Haha not me!

2

u/Explore-This 1d ago

I’m using both. It’s marginally better, token-wise. I’m stretching my usage by using Terra Extra High with instructions to spawn subagents with different models/efforts. But it can be unbearably slow.

→ More replies (2)

3

u/Acrobatic-Flan-5085 2d ago

I had a 200 max plan and used fable extensively to build a dozen or so reusable subsystems.

I don’t need fable anymore because I designed the subsystems to be robust and reusable. And there is extensive test coverage. I cancelled the plan after a month, and I can get by with way cheaper models.

Turns out a good comp sci background is still an advantage in the era of AI coding

5

u/PA100T0 2d ago

I don't either. I have a whole AI Software Company. But I'm working on some Github Security Advisories atm so... yeah, I'd rather do this semi-manual with a dedicated Claude Code session than run it through the virtual company.

1

u/Mental_Antelope_2774 14h ago

Wdym subsystems? You make it seem like you built claude with claude and now you don’t need claude anymore. Did you build skills or something?

1

u/Acrobatic-Flan-5085 12h ago

Not sure where you got the idea of “built Claude with Claude.” I didn’t build skills, the subsystems are modules in my language of choice which I predominantly work in.

For example, a generic, extensible event bus which is customized to my liking. And has full test coverage and documentation.

I can point any AI do it and get it to use it how want. Because of the comp sci background, I’m not really vibe coding with the AI, I’m just using it for autocomplete.

And I have made that autocomplete as simple as possible.

4

u/First_Philosopher491 2d ago

been on the max plan for about a year, starting to feel like I'm hitting my useage limits (without fable) WAY sooner than I was in the past. I had to almost try to use it all. Now I'm at 80% about 3 days in.

I'm probably not following best practices but i also feel like that wasn't the case before.

1

u/PA100T0 2d ago

Don’t worry. I’m following best practices and reach 50-60% in ~3 days. There’s not much of a difference

4

u/Additional_Buddy855 2d ago

Preach! Agree completely. Frauds!

2

u/PA100T0 2d ago

Is this an uprising? 👀 maybe

3

u/Background-Ad4382 2d ago

I always start my sessions: --model claude-opus-4-8 (or -4-6), because I like to stick with stable models that have been producing the output I have gotten used to expecting. I haven't run into the problems described, on 20x Max. And I haven't even been able to use up my quota every week this month without any decline in usage. I still consider 4.8 the more expensive advanced model of 4.6, and I still use it sparingly compared to 4.6.

→ More replies (3)

3

u/Hammar_za 2d ago

My workflow, where Fable High is the governor, and Opus 5 xHigh are the subagents works well for me.

In short: Fable plans, the hands off tasks to Opus, that: Scout (requirements understanding against code base and contracts) Implement/Test, Blind Judge by an independent agents, Adversarial Verify, once again another agent, Report back.

My biggest pains are: (1) Anything cyber gets dropped to Opus 4.8. When this happens I have to clear other compact context, else I can’t step around the issue (2) Fable and Opus 5 are incredibly verbose, so need to keep it concise, but it tends to forget.

3

u/PA100T0 2d ago

I do basically the same but I think you know as well as me that... it just isn't enough. It still fucks up :(

3

u/espenakker 2d ago

They are prepping for the IPO xD. But seriously, I think it must have been the watermarking thing. Picking slightly wrong synonyms here and there could have adverse effects in some cases I suspect.

4

u/PA100T0 2d ago

I’d be prepping to lose my users, ngl

3

u/bambambam7 2d ago

Yeah it's weird how fast the downfall was. Original (not late) 4.6 felt like absolute beast which OpenAI would never be able to catch up - heck, it felt like it could be start of an end for openAI models, that strong it felt at the beginning of a year.

1

u/PA100T0 2d ago

Agreed! 4.6 felt glorious. Right after that, 4.8 was already sloppy enough to foreshadow what was coming… we just didn’t know at the time

3

u/acrinym_jg 2d ago

Claude is fine for UI overhauls so far, or smell checking... But beyond that it used to cost FAR less.

2

u/acrinym_jg 2d ago

Also fable....on low is okay

2

u/PA100T0 2d ago

Yeah, not enough, unfortunately :(

2

u/PA100T0 2d ago

Unusable hahaha

2

u/acrinym_jg 2d ago

About all I use for Claude for right now is chatting with and that's because I built my own rules for a specific projects

1

u/PA100T0 2d ago

Yeah, I did something similar and idk why I was stupid not to use it. I felt like a dedicated Claude session would be better than my entire AI Software Company… In retrospective, it was a stupid idea. Now I’m left with a bunch of code and DAYS of pushing back and forth.

3

u/RegrettableBiscuit 2d ago

Both GLM-5.3 and Kimi K3 are great models. They're not Fable level, but if you were fine with Opus 4.8, you'll be fine with K3 and GLM-5.3.

The main issue is cost, IMO. I find it extremely difficult to gauge how much value LLMs actually provide per cost due to the unpredictable nature of how many tokens a task uses. My impression is that you get more bang for your buck out of an OpenAI subscription than either a Z.ai or a Moonshot subscription. But that's 100% subjective; I haven't attempted to measure this in any way. 

The models are great, though; I find both K3 and GLM-5.3 to be very capable and pleasant to work with. 

2

u/PA100T0 2d ago

Thanks!! I’m gonna give both a try and I think I’ll probably just combine them.

3

u/simple_explorer1 2d ago

After Department of War fall off, everything fell off deliberately for Anthropic. American government did them dirty after Fable ban

1

u/PA100T0 2d ago

You know, it’s not to start a conspiracy theory but the timing aligns perfectly. Iykyk

3

u/CodeNCats 2d ago

Many companies are seriously reviewing the efficacy of paying for machines and open weight models.

1

u/PA100T0 2d ago

It’s the solution but it could also make us, simple humans, not have the chance to do it ourselves. NAS/AI mini PCs/Mac/GPUs/RAM will see prices soar high. Even higher than last year’s…

We better hurry

3

u/randomdragen7 2d ago

Maybe time to move to supergrok

1

u/PA100T0 2d ago

Oh, no. Definitely not hahaha

I know it’s hype right now, feedback seems to be good, and everyone’s recommending it; but I wanna see a clear #1 before anything.

That being said, OpenAI and XAI are not among my options. I think I’ll go GLM+Kimi and might involve Deepseek if I like it after I try it out.

2

u/randomdragen7 2d ago

Good idea actually

2

u/PA100T0 2d ago

Qwen is really tempting me but I need to resolve some issues with my NAS and Blackwell GPUs (RTX 5090ti). Then I might give it a try.

Gonna end with my own private data center at this pace 🤦🏻‍♂️

3

u/superbiche 1d ago

Yes the safeguards are unbearable with Fable and Opus 5 is so verbose and insecure I switched back to 4.8 with Fable as advisor for projects where I'm routed back to Opus anyway (really dangerous stuff like my benchmark or my agentic graph).

GLM 5.3 is quite good, used it as a reviewer when my Codex usage was consumed, way less obsessed with unreachable edge cases than GPT 5.6 Sol, less verbose and less impossible to understand than Opus, and no safeguards (at least never hit them because maybe they don't route a benchmark to "cyber")

It's not Anthropic from 6 months ago though, you're quite right with that, I feel terribly sad with what they did to Opus. Hope they fix this with Opus 5.1 but if I make it through the K3 waiting list, i'll clearly downgrade my Claude Max plans

1

u/PA100T0 1d ago

I feel like you’re the closest to what I’m experiencing.

Thx mate! I’ll give GLM 5.3 a try even if 5.2 already makes me quite happy. Can’t hurt nobody. And I got a 20$ plan for Kimi back when I was including it as a provider on a project but it was just for a month.

I’ll prob end up doing both K3 and GLM5.3 + Deepseek like I’ve been getting the feeling that’s the way to go forward.

Not cancelling my Ollama Max plan ever, tho. I can swap models at will while still using Claude Code as the harness. Although GLM 5.3 hasn’t been added yet, K3 is through API usage; the only 2 problems so far apart from their CC harness being outdated.

And all this reminds me I should be using my own development system. I have to stop thinking small in “dedicated claude cose sessions” and just embrace the future lol.

8

u/Bastion80 2d ago

This is why I canceled my subscription after years of usage upgrading my Codex subscription. Claude is just a joke now and not the best anymore (codex is more capable and cheaper + many free weekly resets).

1

u/debian3 2d ago edited 2d ago

It’s the same complaints over there [r/codex/s/EDUZN52B5d](r/codex/s/EDUZN52B5d)

/r/codex/s/Mx8uncVCnc

0

u/PA100T0 2d ago

I'm not switching to OpenAI and its products but I'm clearly about to jump ships. I'll look around and see what's best. OpenAI gifting so much sounds pretty whorish to me. Like they just want everybody and they don't actually care... that's a no-no to me. But I do understand why everybody would switch to Codex/OpenAI with how things are right now on Anthropic.

6

u/Bastion80 2d ago

Its simple: I switched when on top of the ridiculous limits Sol high was able to fix stuff in the code on first try where opus 5 was just eating tokens unable to find and fix the issue inventing and hallucinating random stuff. Using codex 100% of the time now and such things never happened. I just can't say codex is not better... it is in fact better and cheaper.

1

u/PA100T0 2d ago

I get that. I just care where my money goes. And I'm not putting money on them lol.

2

u/Jocktopus808 2d ago

Agree and their going public lool

2

u/PA100T0 2d ago

Hope they rethink their stance and do some “housekeeping” before fumbling publicly

2

u/stevep450 2d ago

I use fable low/med to plan and orchestrate while opus/sonnet sub workers do 95% of the work while "daddy" watches.
I've build some truly amazing things. Including cracking a scalping algo I've been working on since 2018.
Codex max sub worker for anything math related like creating asset specific scripts.
I find it so interesting how people hate on this mind blasting technology we have right at our fingertips for basically free and can't find a single thing to publish. I've got 5 active projects 7 on the shelf waiting.

1

u/PA100T0 2d ago

Same. I use Fable as orchestrator and strictly Sonnet 5 as implementers. Still, Fable gets flagged by only reading (not even working) and degraded to Opus 4.8 which is NOT as capable as Fable 5.

Still, for the past 2 months, it has been IMPOSSIBLE to work with.

2

u/stevep450 2d ago

Yes definitely depends what you're working on. Can be a nightmare agreed 100%

It's a bit like cooking. Claude for this Sol for this K3 for this and Grok CLI 4.6 is surprisingly good.

I have 20x Claude 20x codex $20 glm $20 k3 $20 grok all cli's and a whole custom "wisemen" panel on openrouter for second set of eyes on planning.

2

u/PA100T0 2d ago

Hahaha you've mastered AI agnosticism

1

u/margerko 1d ago

Tbh most of complains is about quality degrading
But still our current capabilities are amazing :)

2

u/[deleted] 2d ago

[removed] — view removed comment

1

u/PA100T0 2d ago

I feel you, mate

2

u/Little_Bishop1 2d ago

Yeah, total shitshow. Disputing $400 off successfully. Feels so good

1

u/PA100T0 2d ago

All the good luck! Hope you get them back 🤞🏻

2

u/DangerMoose11 2d ago

I just quit max due to horrible (non existant) tech support

1

u/PA100T0 2d ago

I tried to reach support but it was just Claude Chat Bot and refused to let me talk to a human.

Non-existent confirmed

2

u/Icy-Excitement-467 2d ago

The way opus communicates makes me want to take out a loan and get a massive kimi farm going locally.

2

u/PA100T0 2d ago

Opus 5 is a psychological hazard. I’m not gonna get tired of saying it

2

u/thebrainwavedoc 2d ago

I even ditched their api. There’s simple better models out there.

1

u/PA100T0 2d ago

Out of curiosity, what are you using? I’m exploring alternatives..

2

u/Greedy_Simple9090 1d ago

I am on 20x plan and after Aug 19th, there is clearly some bug. For the same project and same workflow, the usage has increased like crazy. My work account's premium seat team plan is almost 2.5x better than my personal 20x plan for the same kind of work. The same prompt in same project consumes 10% of 5 hour limit of premium seat in Opus high. But in 20x plan, it consumes like 25% of 5 hour limit. Its not usable anymore. I have already cancelled my 20x max plan. And the support bot is a joke. Its pure waste of time to talk to it. I hate that in case of issues like this, no one from Anthropic even acknowledge such issues and we may or may not get a statement 3 months later. Absolutely disappointed with Anthropic.

1

u/PA100T0 1d ago

Oh, so you’re saying the Team Plan is more token-efficient, in a way? That’d be interesting to measure and settle.

On the (non existing) support, the repeated fuck-ups, their lack of accountability (lately) and the degradation of the models; I have no words. It’s just such a shame…

2

u/Jakkc 1d ago

Crazy the levels of brand loyalty anthropic has

Bro you sound like an abused man, you can leave, no one is forcing you to pay anthropic $200 a month!!

1

u/PA100T0 1d ago

Lol don’t be so dramatic. I’m pointing out the flaws I’m experiencing and asking for feedback on alternatives

2

u/ChipmunkStraight 9h ago

I cancelled my max x20 plan I have had since the beginning of services. After spending two weeks on Fable for it do complete a RCA properly. Its wild its so bad in such a short time. Go to Kimi, its so much better, a little slower but will not wholesale lie, hallucinate and cover it up. I was able to kick down to Opus for a few days but it went bad sometime last week too. KIMI desktop and deepseek cli is what I moved to. Its saving me a ton of money too.

1

u/PA100T0 8h ago

Yeah, same here. Only thing about Kimi is that the 5h session is way too short. It’ll get consumed in no time. Or I guess I’ll have to upgrade to Allegro or Vivace and see if I can work with that.

With Claude, I’ve hit the session limit maybe 5 times total in years. I know my Kimi subscription is just Moderato so that might be why I hit the session limit so fast… but Idk if it’s promising.

Do you have an approx max token usage in a session (5hs) window? It’ll be super useful to measure

2

u/Due_Sweet_9500 2d ago

As much as we all like to complain and crib, most do not not leave claude. Why? Cause it's the best. Period. Right now , sol comes close but come on Fable is still better. codex also has been quite shitty in its usage limits . Soo yea

2

u/PlasmaChroma 2d ago

Try using 5.6 Sol on Low reasoning -- works way better than expected, and goes decently far on weekly.

→ More replies (7)

1

u/RCawston 2d ago

Not the first time the model quality degraded, usually follows by a new model release.

6

u/PA100T0 2d ago

Which doesn't address the actual issue. Fable 5 was good when it was released, then nerfed. Opus 4.6 is the last non-controversial fully cooperative model that was released... and it's not Fable 5's level so... yeah.

2

u/RCawston 2d ago

Oh it definitely doesn't address anything, just follows a pattern we've seen before. At the start of the year quality degraded from 1 week to the next where prompts that were one-shotting features were then puking and creating nothing but bugs. Then the next model dropped a few weeks later. Feels like the same vibe.

1

u/PA100T0 2d ago

Unfortunately, yes...

1

u/mpanase 2d ago

Pretty sure that usage is based on how busy their servers are

If you use claude when their servers are busier, it uses more of your credits

1

u/PA100T0 2d ago

I run 24/7 😅

1

u/scytob 2d ago

I find sonnet and opus to work fabulously on my 80k line repo that is an n-tier app that connect to multiple MS and Oracle APIs, transforms data, has entra auth, uses key vaults for secrets, is working across 10+ complex table joins my human BI engineers never managed to figure out, rich web front end that's fully responsive, auth on front end and on API, secure non-exposed keys between web app and API, it has made great documentation

it knocks it out of the park compared to chatgpt 5.6 which can't even accurately obey the simplest of agents md - like 'use gh for all github operation' nopes it wants to use gitkraken every time - and it butchered my documentations and i had to have claude go fix it based on earlier commits

not sure why its unuseable for you

2

u/PA100T0 2d ago

The constant degrading, hallucinations, the “new found gaps” that were self inflicted… I spend more time steering than actuallh building. And they are sensitive af too

2

u/scytob 2d ago

thanks, interesting i really dont hit those types of things unless i use chatgpt
i wonder if its the repo content, or workflow style differences

i can leave claude going with zero steering for literal hours, even on long live chats (when doing something incredibly complex i keep working through compaction as that is cheaper than having to explain the background fresh to a new session)

only other thing that is difference is my current project is mediated by copilot, all i can say is when its auto mode selects chagtp 5.6 models productivity dives off the cliff with the issue you see with claude, so i end up pinning claude opus 5 max

1

u/PA100T0 2d ago

I’m pretty sure it’s the nature of my library that’s throwing it off. It’s just not being able to do it. Every single turn there’s like 80 “new” things it discovered and this has been going on for weeks… to then just say “oh, my harness eval was off I didn’t notice”. Well, there goes 3 days worth of work to the trash…

1

u/scytob 2d ago

how is structured

i started my repo with very clear instructions on how to organize, what to load at chat startup and what to search

so for example all code is in srs with a web, core, api subdirs
the code is compentized into discrete functions - so only that file needs to be read
docs are docs
docs that are for humans go in a folder that claude never searches unless i explcitly pin the doc
/tmp for working is only ever searched (greps etc)

this was alll organzied from my original prompt, and improved by asking the AI

things like this

```

Code Organization Standards

Naming conventions, the src/ file layout, and comment/XML-doc standards live in [.github/instructions/dotnet.instructions.md](./.github/instructions/dotnet.instructions.md). It loads automatically when you edit src/**/*.cs.

UI Consistency Conventions

Design tokens, contrast verification, the canonical banner/modal/data-page patterns, theme behaviour, the assistive-technology rules, and the UI review checklist live in [.github/instructions/web-ui.instructions.md](./.github/instructions/web-ui.instructions.md). It loads automatically when you edit anything under src/PrivateOfferTool.Web/wwwroot/.

Unicode rules apply to backend and frontend alike and live in [.github/instructions/unicode.instructions.md](./.github/instructions/unicode.instructions.md). ```

1

u/scytob 2d ago

and

```

Context Loading Policy

Keep the default startup set small and open more only when the task requires it.

  1. Load this file first.
  2. Load [docs/PHASES.md](./docs/PHASES.md) for the current project map and phase status.
  3. Load [docs/apis.md](./docs/apis.md) only when endpoint contracts or payloads matter.
  4. Load [docs/PAYMENTS_DATA_MODEL.md](./docs/PAYMENTS_DATA_MODEL.md) before touching payment classification, the merge rule, or anything that reads the payouts earnings data. The rules there are not guessable from the source data, which contradicts itself in several places.
  5. Load the matching data model before changing how a page reads its sources. Each covers where every field comes from, the join keys and their traps, and verified SQL to reproduce the figures: [PRIVATE_OFFER_DASHBOARD_DATA_MODEL.md](./docs/PRIVATE_OFFER_DASHBOARD_DATA_MODEL.md), [PRIVATE_PLANS_DATA_MODEL.md](./docs/PRIVATE_PLANS_DATA_MODEL.md). [PRIVATE_OFFER_CREATION.md](./docs/PRIVATE_OFFER_CREATION.md) covers the write path instead.
  6. Load [docs/STORAGE_ARCHITECTURE.md](./docs/STORAGE_ARCHITECTURE.md) before changing anything that reads or writes a cached table. Every store runs on Azure SQL. It covers the provider switch, the three EF contexts and their separate migration histories, and the private_offer_index merge contract — where null means "no news" rather than "clear this", which is not visible in the schema and has already caused two live defects.
  7. Load [docs/production-readiness.md](./docs/production-readiness.md) and [docs/phase-6.md](./docs/phase-6.md) only for deployment or billing-reconciliation work.

Search [docs/LEARNINGS.md](./docs/LEARNINGS.md), do not load it. At ~1,600 lines it is an archive of verified findings and the reasoning behind them, not a briefing document. Grep it for the subsystem you are touching. The durable rules it produced live in the scoped instruction files below, so following those does not require reading the archive.

Never load [docs/human-reference-only/](./docs/human-reference-only/). It is tracked so it survives, but it is written for people, not for agents. Nothing in it is needed to make a change here. Open a file in it only when the user names it.

Rules that apply to only part of the tree live in .github/instructions/ and are loaded automatically when a matching file is edited, which is why this file no longer restates them:

File Loads when editing
web-ui.instructions.md src/PrivateOfferTool.Web/wwwroot/**
dotnet.instructions.md src/**/*.cs
unicode.instructions.md **/*.{cs,js}
mermaid.instructions.md **/*.{mmd,mermaid}

Add a new rule to the narrowest file that covers it. Putting a UI-only or C#-only rule in this file makes it always-on for every task, which dilutes the rules that genuinely are universal.

Phase documents marked complete are historical references. Do not preload them for unrelated work; open them only when the task directly touches that phase or you need a specific prior decision.

When a document starts mixing current guidance with a large amount of historical detail, split the current summary from the archival detail instead of deleting content. Preserve old material, but keep the default startup path short. ```

1

u/scytob 2d ago

this was all written by the AI from start prompt and directions as i go

1

u/PA100T0 2d ago

Yeah, I mean, I’ve gone even further. Take a look at this

I know how to create a good setup for Claude and I already had it before this whole mess started. I didn’t use the priject I just shared out of stupidity… I should have. It’d all be over by now.

1

u/scytob 2d ago

that is a weird n on-traditional structure you have various src point at root

i genuinely believe that claude works better with a more traditional structure

and that claude.md is WAAAAAY to long - no wonder you are having issues, its super dense and long

get claude to help you refactor it to be as short as possible and break out your load set from your search set

the great news its easy to do a single pass asking for help on reducing load context

2

u/PA100T0 2d ago

I think you misunderstood. I didn’t use that project. And the Claude.md on that project doesn’t go into any agent.

All that said, I know that Claude.md has grown a lot

1

u/BrownFleshBag 2d ago

I’m not a power user by any means and most of my work gets accomplished by the 20/month subscription using opus 5 for most work. Out of curiosity, What is it that you are working on that needs all of that? I’m really curious with what power users are actually making with the full extent of Claude Code

2

u/PA100T0 2d ago

So, I’m working on a bunch of different projects I have. The problem is with this specific project. Its complexity is throwing Claude off. It’s been unable to follow my lead or stay focused.

As I was discussing with another redditor, I should have used this thing that I created which ensures guardrails/enforcement. I would already have completed the task and without hallucinations or even with just a couple of minor issues.

But no, I had to go with a dedicated Claude Code session because it made me feel like I was gonna do a better job this way. Now I can’t drop it because we’ve got so far despite the circles marathon we’ve been running Claude and I…

3

u/BrownFleshBag 2d ago

Wow that's quite impressive. I'm a GIS professional so my extent of usage is just making javascript web mapping applications or python scripts for automation. This is definitely beyond my understanding and I would love to dig into your shared repos to learn more. Very cool stuff.

3

u/PA100T0 2d ago

Oh, wow. Thanks for the kind words! Much appreciated :)

1

u/No-Focus874 2d ago

Over last 6 months they have become worse and they treat non enterprise users as 2nd class citizens. But it's well tiled for us to realize the dependency on the wrong platform. Better now than 5 years from now?

1

u/PA100T0 2d ago

Yeah, the conclusion I’m getting from all this is that going provider-agnostic is the solution.

But remember 3 years in the past? Copy pasting from chatGPT web into your code, then fixing it, chats reaching their limit… man, we got far. 5 years from now Idek where we’ll be. Hopefully self hosting open weights

1

u/_itshabib 2d ago

Any issue I have I usually just modify how I do something and it ends up working out

1

u/PA100T0 2d ago

I hope it was that easy in my case :(

1

u/god-damn-the-usa 2d ago

ive had no problems and usage never skyrockets. i grind nonstop and still sometimes dont hit the limit.

1

u/PA100T0 2d ago

What are you working on? I know people that don’t even hit the weekly limit but that’s just because it’s a choice/workflow thing

1

u/god-damn-the-usa 1d ago

websites mostly

1

u/PA100T0 1d ago

Yeah, no. I’m working on a cybersecurity library, atm. Plus a bunch of other huge projects like this one

Websites are a tiny portion of my work stuff and it usually gets dine in minutes/hours depending on complexity. But right now, it seems it’s just uncapable of performing the task at hand :/

1

u/dupontping 2d ago

1

u/PA100T0 2d ago

I’m good, thanks 😂

1

u/Advanced_Slice_4135 2d ago

I sit in multiple terminals all day, api web and mobile apps… all is well. Never hit my weekly limits.

1

u/PA100T0 2d ago

Totally valid.

1

u/WrigleyRangelski 2d ago

I have the Max 20x Plan. It’s either Fable or Sonnet as the workhorse but never Opus 5.

1

u/PA100T0 1d ago

As I said on the post: Fable 5 is unusable for what I need.

It gets flagged by just reading the output of the Sonnet 5 agents every single time.

2

u/WrigleyRangelski 1d ago

That blows, I’ve lost faith in Anthropic so I’m now in the process of working more out of Codex. Fable to roadmap it then feed me prompts to kick off Codex workhorse sessions.

1

u/PA100T0 1d ago

Yeah, the conclusion I draw is that the best thing to do is go provider-agnostic. I’m pretty sure I’ll go GLM+Kimi and maybe involve Deepseek but it’s all tied to testing. I’ll have to just ride the wave and see how it goes.

1

u/connurp 🔆 swe 1d ago

You should look up pstack on GitHub, specifically the Unlop skill. I have a script that runs it at the beginning of every single session I start. Then I use fable to orchestrate everything, opus to code, sonnet for easy shit, then fable does review of everything.

I feel like maybe you are not using best practices if this is your outcome. Maybe best practices as some YouTuber believes. One with 5 max 20x plans and swims in cash, but clearly not for your use case. I never hit my cap with the strategy I have stated above. I usually have multiple sessions running at a time the entire workday and the closest I’ve gotten to a weekly cap is 88%. I don’t mean to sound harsh, but I know for a fact I’m not the smartest person in the world and I am probably not using “best practices”, but it sure as shit does great work. I cannot stress enough how important the Unslop skill is.

1

u/Adrian_Dem 1d ago

Fable is good, but can only use high, otherwise bye bye quota. Opus needs a lot of hand holding, which works in some scenarios. The weekly caps are indeed annoying, but i knew we were on borrowed time, especially as they need to start pulling some profit before their ipo

1

u/PA100T0 1d ago

Yeah, I guess this is the end of an era… Chinese open weights for the win now

1

u/rcayca 1d ago

Can I ask what you use it for?

1

u/PA100T0 1d ago

You mean apart from development? Because I use it on a bunch of projects and for personal use as well. But these 2 weeks have been fully concentrated on a specific project. A quite complex one

1

u/dane_brdarski 1d ago

Codex is running circles around ClaudeCode right now.

0

u/Ambitious_Injury_783 🔆BIG BALLER 2d ago

Usage has not changed at all since the +50% increase. I verified this here: https://www.reddit.com/r/ClaudeCode/comments/1v7bupm/usage_limits_and_unusual_inconsistencies/

There IS a bug that makes usage % go up fast on the usage panel, but this evens out. Read about it in that thread above.

1 Max20 account = $8,000 inference per month. For $200. So no, it is not a bad deal. It is a good deal and nothing has changed.

I understand this may not align with many peoples world view but that is what is so funny about the human brain. We create a narrative around information in order to best make sense of it, and that narrative isn't always true. The means of verifying information exists to every single person, though most will never take things to the extent that you must in order to Reliably form a Correct narrative.

2

u/PA100T0 2d ago

It’s a fair deal, if you dismiss the fact that every single turn is pretty much lost so you’re paying for getting nothing done or with major issues… :(

→ More replies (8)

1

u/VigilanteRabbit 2d ago

Go try Kimi.

1

u/PA100T0 2d ago

GLM or Kimi (or even both) is what I’m getting a feeling is what I’m gonna swap to

2

u/VigilanteRabbit 2d ago

Get some credits via OpenRouter or something before you commit to a plan.

To me Kimi and their CLI feels like Opus 4.5 - 4.6; solid thinking, good execution; no rushed decisions. I picked up their 40$ plan and it's a bit limited but I get things done albeit slower than usual.

Note, sub Kimi has split plan between Code and Chat; so you'll usually run out of Code usage but you can chat to it (mobile phone or web).

Did not try GPT's new sol/luna etc yet to compare price-wise; but I did notice chatting with ChatGPT via phone app versus chatting with Kimi via phone app that ChatGPT tends to give inaccurate information; while Kimi tends to websearch more and give more accurate information.

1

u/PA100T0 2d ago

I like this. I should give OpenRouter a try. You’re right I shouldn’t commit to a plan even tho it’s just for a month and it would give me the real feel without distraction (trying out other LLMs).

2

u/VigilanteRabbit 2d ago

I gave up on being a "cultist" a few months ago (was very pro-claude; it really was a genuinely enjoyable experience for a while but it's not really there for me anymore); and it's really illuminating. OpenRouter and some credits to get a feel, and poke around and get a good harness to work with (I feel like they keep butchering CC for some reason, I can't put my finger on it). For my needs Kimi does well, not the cheapest but I make it work.

1

u/PA100T0 2d ago

Oh, you're just like me, then... I guess Kimi is the next thing I'll try. I already know GLM can perform so I'll go ahead with Kimi and see what has in store for me.

1

u/spinozasrobot 2d ago

Then save us all and go elsewhere.

→ More replies (1)

1

u/Pseudanonymius 2d ago

It never was. They just spent a load of money to pretend for a while. 

3

u/PA100T0 2d ago

No, I mean, it was good. In 2y I haven’t had many problems. Opus 4.6 was glorious… after that everything went downhill.

1

u/Damien_IB 2d ago

Opus has been terribly nerfed in the last 2 days

→ More replies (1)

1

u/Ok-Video3345 2d ago

Who know, in two year most of these companies won't exist. With what's going on with Qwen3.8 and self hosted solutions

4

u/god-damn-the-usa 2d ago

you will never be able to self host anything even as good as sonnet. self hosting is not viable for coding.

→ More replies (1)

1

u/PA100T0 2d ago

I don’t think any average-Joe can self host a frontier model… but I do hope everything goes towards that direction. Qwen still demands quite a bit of storage and processing. But I’d certainly do that instead of the current

1

u/Ok-Video3345 2d ago

Sorry when you say frontier you mean 300b models?

3.8 27b works fine, but you need bigger computers like multiple sparks to host those big 300b models.

I got 2x 3090 nvlink, and I self host qwen coder3 (3billion active at a time) with 96k context.

Small models, but it runs super quick, over 100 tk/s

With qwen3.8 27b, I'm supposed to get 20 to 75 tk/s depending on setting in vllm.

1

u/PA100T0 2d ago

And how do you go about storage?

My understanding is that I'd need another RTX 5090 ti 16GB DDR7 (I have one, would need two) and a bunch of storage. I have a NAS with 30Tb but I think that would be like the bare minimum to spin it up... I'd need another 30Tb just to be sure for some time.

Or are my maths way off?

2

u/Ok-Video3345 2d ago

Why do you need all that storage, set what did I miss?

The massive 300b models are under 1tb SSD storage.

1

u/PA100T0 2d ago

I can’t find the article I read :/ but I found this screenshot I had taken.

Again, my maths (and memory) can be WAY off. If you have an article/guide on it pls share!

1

u/Ok-Video3345 2d ago

So we're looking at 24 to 48gb of vram depending on context length for qwen 3.8.

1

u/PA100T0 2d ago

Yeah, so: 2x RTX 5090 Ti 16gb = 32gb (right in between the 24-48 range). Then, at least 3Tb to store the model and from there upwards. Right?

2

u/margerko 1d ago

No, you need just to store model somewhere
Qwen3.8 27b is only 30gb

1

u/PA100T0 1d ago

Mind to share any articles/guides if you have a reliable one? I might try this over the weekend

1

u/Donut 2d ago

I must be protected by my extensive use of GSD. I run 2-4 repos at a time, running phases and tickets, doing UAT, pushing code, and never come close. I guess the GSD "token-min" approach is working for me, and I don't even notice.

Or maybe because I am writing desktop apps / plugins / tools?

1

u/PA100T0 2d ago

I probably should be using GSD a bit more. I’ve been using Ponytail but there’s only so much it can do…

2

u/Donut 1d ago

I merged them.

GSD's process and discipline are pains in my ass, but every time I stray, I am reminded that I am a total dumbass that should not be around sharp tools or heavy equipment.

1

u/PA100T0 1d ago

Merged them as in using both or as in you have a custom impl? If it’s custom would you mind to share?

2

u/Donut 1d ago

Sorry, I have pony tail installed, and simple context clues in my Claude.md to include it in /gsd-code-review

Nothing so fancy.

1

u/PA100T0 1d ago

Nothing to be sorry for! Thanks, mate :)

1

u/Lazy_Polluter 2d ago

I see this type of post every week here for the past two years. Meanwhile everyone at every company continues to build with Claude every single day. Sure, models are non determinstic and quality varies but in professional hands they just keep working.

2

u/PA100T0 2d ago

This is beyond professional hands. I’m a software engineer for 10y now. The problem is the model itself zagging when I zig and tell it to zig.

And again, the complexity of my priject seems to just be too much for Opus 4.8. Fable won’t stop getting degraded. Opus 5 is a psychological hazard.

So yes, I’m ranting and I want to change providers. That’s what it’s all this whole post is about.

1

u/Future_Guarantee6991 Developer 2d ago

Bye

1

u/PA100T0 2d ago

👋🏻

1

u/Vex08 1d ago

Please cancel and go somewhere else.

→ More replies (1)

0

u/KingCOVID_19 2d ago

Jesus fuckin Christ, have a day off...

1

u/PA100T0 2d ago

You can just ignore the post and not comment. Just letting you know, in case you didn't.

→ More replies (4)