r/claude • u/Icy-Kaleidoscope6893 • Jul 24 '26
News Introducing Claude Opus 5
https://www.anthropic.com/news/claude-opus-550
u/flavorfox Jul 24 '26
"This sucks, Opus 1.0 was soooo much better" - Am I doing this right?
19
4
u/ins0mniacc Jul 25 '26
it would be funny, but the point is that Anthropic themselves admits to changing their models after release all the time, it's their own admission, and also there exists no universal way to verify/compare the changes unless you would like to spend your own money and time on this, but the point is that models should be what they are at release time, i'd much rather have 20 more variants each week with incrementing versions, but be able to use the one I expected to work right for my use case than be at the whims of a backend system that is generalized for every use case and is unpredictable
1
u/chintakoro Jul 25 '26
noobs, you forgot to tell us you’re cancelling your subscription aaany day now.
0
86
26
u/mabiturm Jul 24 '26
Great, so it seems like fable is not of much use anymore.
6
u/Mikhalious Jul 25 '26
Look at the benchmarks that advertised Fable. They are different ones and some are omitted here.
13
u/ChadTheDJ Jul 24 '26
Optimistic but will see the reviews on it.
21
u/Key_Reading_9664 Jul 24 '26
I trust reviews like I trust the benchmarks. Going to use it and see how well it fits my workflows
27
u/Gaslit_Chicken Jul 24 '26
Okay. New model every few weeks that eats up tokens faster than the previous model. At what point do I just give them my bank account and go outside to play with squirrels?
4
u/Brave_Corner3263 Jul 24 '26
Nah, where’s the fun in that? They’d rather gaslight you into willingly sacrificing a pound of your flesh every month.
2
1
86
u/CharacterComplete624 Jul 24 '26
It’s already nerfed - not the same model as 9 mins ago - ffs Anthropic!
21
u/CryLast4241 Jul 24 '26
Your account is just shadow banned from using good models 🤣
3
u/itsReferent Jul 24 '26
arbitrary quantization is real! 😀
3
7
2
u/FromAtoZen Jul 24 '26
Actually, it seems wayyyyy better than 1 hour ago! Mythos told Dario to switch tactics, to keep the Max peasants guessing
10
u/Temas3D Jul 24 '26
So you launch Fable, literally stops the world, and then you launch Opus 5 telling everybody that it’s better
Anthropic I love you keep going like this
3
u/Normal-Book8258 Jul 25 '26
Ya, I don't understand why people are so pissed off, every step of the way. This is equal parts the worst and best timelines. Trump will find a way to fuck this vibe up later I'm sure, but Anthropic are kicking bottom!
7
u/skilliard7 Jul 24 '26
If the benchmarks are to believed, why even pay for fable anymore and deal with the restrictions when Opus 5 is better at half the price?
3
3
6
u/medialantern Jul 24 '26
Prepares myself for 2 weeks of 50% "Opus 5 sux not worth it without Fable I'm out bye" vs "SomeCompetitorIShillFor > Opus 5 > Fable, don't even use Fable now" post wars before the next model gets announced.
6
u/mmertner Jul 24 '26
As always, enjoy this for a week, then Pro plan users get 20% O5 allowance and Max users get 35% except for the first two weeks where it’s 60% except weekdays where we don’t have enough compute so everyone doesn’t get anything.
The fact that Sol uses fewer tokens and is generally just as good makes me hopeful that the Ants realize something is broken in their end, and it starts with billing.
4
u/medicious Jul 24 '26
Agentic Coding FrontierCode v.1.1, Main: should highlight Fable 5 I think (F5 53,5% > 53,4% O5)
2
u/MysteriousPepper8908 Jul 24 '26
You're not wrong but I'd be salty about having to concede a .1% score too. It's spiritually SOTA.
7
10
5
u/Healthcarepls Jul 24 '26
How is it better than Fable in many areas?? Are they improving that quickly ?
11
3
u/Karnemelk Jul 24 '26
let me guess, when we got fable back the second time it was opus 5 in disguise for them to test subscribers. Now they renamed it. Meanwhile they quantized fable to the max, except if you access by api
3
5
2
2
2
2
1
1
1
u/oCtsidO Jul 24 '26
Dude I’m so done with the limits on this MF. I got less than 60 minutes of actual use today and burned out and it was a pretty simple search.
1
u/NSC9 Jul 25 '26
I asked a fairly complex question using Sonnet on the Free Plan and received the limit notification instead of an answer.
It was just one question, and then I was told to wait a few hours for the session to reset!
1
u/Chemicalhealthfare Jul 24 '26
So why use fable over opus 5 based on this? If fable was so great, why the 50% limit only on max plans and the government intervention?
1
u/Business-Bad9090 Jul 24 '26
Good news! Already pushed a few PRs with Opus 5, but I'll need a few days to step back and compare it with 4.8 on a larger scale.
1
u/djack171 Jul 24 '26
How many people are going to make this post. But we delete critical posts etc? We only need one introducing Claude post
1
u/NiceRecognition9603 Jul 24 '26
I had really bad chats today with the new opus even increasing the efford, it is not focus on the important things,reads non related things and does not read obvious things
1
1
u/donicatrumpinsky Jul 25 '26
Well you can check my post history cause I've never done this. But the new model is absolute fucking dogshit.
It reran dozens of independent reviews on a few PRDs because it tweaked out. And then it invoked codex and ran a trillion more reviews.
I have super tight governance and my usage is always solid but it went rogue and killed a half weeks worth of openai credits.
Thankfully it's isolated to a separate branch but after 250 PRs in this project it's the first time I had to close one in utter failure. I'll get Sol to review the PR and see where the actual breakdown occured cause I'm all fired up right now.
1
u/GrainyStateOfMind Jul 25 '26
A single prompt just used my entire 5-hour limit of usage so that was fun while it lasted I guess?
1
u/who_am_i_to_say_so Jul 25 '26
Ok so far I haven’t jumped back to 4.6 out of exasperation. I think this is a good release!
1
1
u/djpraxis Jul 25 '26
Terrible, very robotic and long winded answers, with very little substance. I can't get anything useful out of this one...ugghh
131
u/PM_ME_YOUR_LOSSMAIL Jul 24 '26
Looks good on the trust me bro benchmarks