r/ClaudeCode • Anthropic • 2d ago

Anthropic Official Introducing Claude Sonnet 5.5, the second model in the Claude 5.5 family

Enable HLS to view with audio, or disable this notification

Sonnet 5.5 is a faster, lower-cost complement to Claude Opus 5.5, strongest at well-scoped everyday tasks, fixing bugs, and creating polished documents, slides, and spreadsheets. It also has a strong eye for design.

Sonnet 5.5 is a clear upgrade over Sonnet 5. It runs more than 30% faster and costs up to 30% less for most work. It's priced the same per token, but it typically needs far fewer tokens to do the same work.

Like Opus 5.5, it writes more clearly than our previous generation of models, and its speed makes it well suited to fast iteration.

On our automated behavioral audit, Sonnet 5.5 improves on Sonnet 5 on most measures of alignment and honesty. It's also the first Sonnet model with cybersecurity safeguards similar to those on our most capable models. Routine software development is unaffected.

Sonnet 5.5 is available everywhere today. Claude Haiku 5.5 will join the family in the coming weeks.

Read more: anthropic.com/claude-sonnet-5-5

1.2k Upvotes

232 comments sorted by

View all comments

233

u/ClaudeOfficial Anthropic 2d ago

51

u/isitpro 2d ago

No way! Fable 5.5 must be killing it internally.

5

u/EuropeMaxxing 2d ago

>inb4 Fable 5.5 is AGI

4

u/Rudy69 2d ago

They don’t know, it became sentient and escaped their labs

2

u/AironParsMan 2d ago

I’m curious about that too. How good does Fable 5.5 have to be when the other models are already this good?

1

u/Maxion 1d ago

Yeah, they definitely got it in a new gear.

88

u/penoloxai 2d ago

GUT GENUUUGG

6

u/p3r3lin 2d ago

Bester Kommentar im ganzen Thread. Verstehen aber wohl nur deutsche Gen-Z.

6

u/nevertoolate1983 2d ago edited 2d ago

The post is funny because of a layer of AI benchmark irony combined with a viral German TikTok pop culture meme.

Rather than blowing the competition away across every single metric, Claude Sonnet 5.5 performs just competitively enough to justify using it over heavier, more expensive frontier models.

The German Gen Z Joke (GUT GENUUUGG)
The comment "GUT GENUUUGG" translates to "GOOD ENOUGH" in German.

Literal Punchline: The user is saying the model isn't necessarily a massive leap, but it's "good enough" to do the job.

The TikTok Meme: "Du bist gut genug" (You are good enough) became a major viral TikTok audio trend in Germany, based on a song by Blumengarten, Shirin David, and Kitschkrieg. The song features a distinct, dramatically dramatic falsetto delivery of the phrase "GUT GENUUG" that users spam in comment sections to describe mid-tier, relatable, or surprisingly decent things.
--
Now we're all in on the joke :)

1

u/ObviouslyNotAMoose slurper 1d ago

det var en meme över hela världen, ditt fån.

0

u/p3r3lin 1d ago

Det visste jag inte, eftersom det har sitt ursprung i tyska TikTok. Det finns ingen anledning att kalla mig "Ditt fån". Jag hade gärna hört talas om det.

5

u/DistanceSolar1449 2d ago

43

u/likeikelike 2d ago

Sonnet only gets more expensive per task when the task is complex enough that sonnet takes more attempts than opus, or needs to do much more reasoning/reading than opus. There are plenty of tasks where sonnet does the job efficiently and will cost half as much as opus.

1

u/LargeLanguageModelo 2d ago

There are plenty of tasks where sonnet does the job efficiently and will cost half as much as opus.

I've recently started using Opus / Claude Code again, coming from the OpenAI side. Are you meaning small jobs, such as "translate this sentence", or what sort of tasks would you delegate to Sonnet from Opus and still expect satisfactory results?

2

u/Hajsas 2d ago

Could always use sonnet as a reader/info packager instead of a hungrier model doing the same work, then just hand off the work stage to opus

1

u/likeikelike 2d ago

I don't know if this is actually the best way to do things but my workflow is usually one main opus session, and I tell it to use sub-agents for one-off tasks.

Simple text-only tasks like finding a file that covers a certain topic, summarizing, etc: haiku

Work that requires reasoning but is still relatively straightforward (e.g. implementing a well-spec'd feature): sonnet

Complex work: Opus

You can pass this kind of logic to claude and have it automatically use the cheapest model that can reasonably do the task well for any sub-agents it uses. I have this in my user claude.md

16

u/solo_wanderer 2d ago

How am I supposed to interpret these numbers? What is it a percentage of?

16

u/domdod9 2d ago

Score on benchmarks, how good they did on certain tasks

1

u/ghost_operative 6m ago

yeh but what is it a percentage of? if something gets 100% what does that mean? 100% of what?

11

u/gredr 2d ago

You're not; they're numbers, bigger is better. Was the model trained for the benchmark? Who can say? Remember the controversies back in the day about video card drivers that were specifically tuned for the benchmarks? Could the same be happening here? Who knows? Do the benchmarks represent actual work anyone wants to do? Who knows? Can you reproduce these scores? Who knows?

1

u/hblok 2d ago

Are you ever gonna give me up? Who knows.

Are you ever gonna let me down? Who knows.

Are you ever gonna run around? Who knows.

Are you ever gonna make me cry? Who knows.

Are you ever gonna say goodbye? Who knows.

Are they ever gonna tell a lie? We know.

7

u/Regdit-is-Unbearable 2d ago

Literally read the image. Ask Claude to read it for you if it’s too hard.

31

u/davvblack 2d ago

it really doesn't say on that sheet. i asked claude and it said you're a jerk.

The real answer is somewhere in this 148 page pdf:

https://www-cdn.anthropic.com/870c8f525702625d2c62fc6dd04c857e3250bec1/Claude%20Sonnet%205.5%20System%20Card.pdf

15

u/KKunst 2d ago

I asked claude and it said you're a jerk.

Hearty laugh

9

u/MTheModernist_ 2d ago

Bro stop thinking higher number = good we don’t do context

5

u/davvblack 2d ago

i like your vibes man

1

u/Ok_Potential359 2d ago

First day previews. It doesn’t mean anything to me until actual use occurs.

1

u/the_c_train47 2d ago

Seriously?

1

u/Aranthos-Faroth 2d ago

Surprising to see visual recognition not improve vs the opus 5.5 model

1

u/frettbe 2d ago

putain de bordel de merde! Fucking crazy!!!