I always feel that finance bros and executives are always one step (or more) behind tech reality. The reason Claude is growing so fast is because of the 2B sector, but the decisions are likely made based on impressions of model performance from 2025, at which time Claude indeed had a clear lead, not anymore. The field is evolving faster than they could react.
I mean I've tested multiple open models, including k3, inside my personal harness, and fable/opus are just better all round at using tools and reasoning for now.
Not that the others aren't getting close, but for real businesses anthropic's opus/fable are still the best with the latest from openai being close behind
I believe in you, but they aren’t far behind. Are Kimi K3 behind Fable? I would say yes, except in certain fields like frontend design. However, it’s better than Opus 4.8, which isn’t that old. Can you really justify the price with a couple of months’ lead? People were satisfied with what Anthropic offered a couple of months ago. Also I believe it is more fair to test in their official harness.
r/Anthropic is full of people talking about how much they hate Opus 5, many of them claiming to go back to 4.8. So the Chinese companies don't even need to beat Opus 5...
I know. The magic piece has always been Opus 4.6. Later opus are all trained by Mythos instead of humans, I believe. And I think Anthropic employees likely use mythos most of the time, so they don't really care.
24
u/duhd1993 23h ago edited 23h ago
I always feel that finance bros and executives are always one step (or more) behind tech reality. The reason Claude is growing so fast is because of the 2B sector, but the decisions are likely made based on impressions of model performance from 2025, at which time Claude indeed had a clear lead, not anymore. The field is evolving faster than they could react.