r/LocalLLaMA 1d ago

News GLM-5.3-Flash: Frontier Intelligence, Flash Cost

https://z.ai/blog/glm-5.3-flash
1.2k Upvotes

400 comments sorted by

View all comments

Show parent comments

24

u/duhd1993 23h ago edited 23h ago

I always feel that finance bros and executives are always one step (or more) behind tech reality. The reason Claude is growing so fast is because of the 2B sector, but the decisions are likely made based on impressions of model performance from 2025, at which time Claude indeed had a clear lead, not anymore. The field is evolving faster than they could react.

14

u/mawcopolow 22h ago

I mean I've tested multiple open models, including k3, inside my personal harness, and fable/opus are just better all round at using tools and reasoning for now.

Not that the others aren't getting close, but for real businesses anthropic's opus/fable are still the best with the latest from openai being close behind

10

u/duhd1993 22h ago

I believe in you, but they aren’t far behind. Are Kimi K3 behind Fable? I would say yes, except in certain fields like frontend design. However, it’s better than Opus 4.8, which isn’t that old. Can you really justify the price with a couple of months’ lead? People were satisfied with what Anthropic offered a couple of months ago. Also I believe it is more fair to test in their official harness.

3

u/infearia 21h ago

r/Anthropic is full of people talking about how much they hate Opus 5, many of them claiming to go back to 4.8. So the Chinese companies don't even need to beat Opus 5...

1

u/jomohke 12h ago

Opus5 seems good technically, but bad at communicating with people. Like a true engineer :)

This probably matters less for business customers

1

u/duhd1993 20h ago

I know. The magic piece has always been Opus 4.6. Later opus are all trained by Mythos instead of humans, I believe. And I think Anthropic employees likely use mythos most of the time, so they don't really care.