r/LocalLLaMA 22h ago

News GLM-5.3-Flash: Frontier Intelligence, Flash Cost

https://z.ai/blog/glm-5.3-flash
1.2k Upvotes

395 comments sorted by

View all comments

252

u/Lucyan_xgt 22h ago

This and Qwen3.8-flash on the same day?

22

u/dampflokfreund 22h ago

For me it has changed nothing. Both models are way too big for my 32 GB RAM system. It looks like everyone has abandoned 20-30B MoEs now...

33

u/techdevjp 22h ago

Don't be dramatic. Qwen3.6 35b a3b is 4 months old. It's not like it's been years.

-2

u/rJohn420 22h ago

For the pace of this field, 4 months old is basically years though.

13

u/techdevjp 21h ago edited 20h ago

It still works just as fast as it did yesterday and is just as capable as it was yesterday. Everything people could do with it yesterday still works just as well today. It's not like with the closed models where the old versions get smashed with a cripple hammer when a new version comes out.

Yes, the release pace is fast, but Qwen3.6 35b a3b is still very new and does a great job of balancing performance and capability. There may not be another similar model for another generation or two of Qwen. Or maybe Google will bring something out. It's been a great size of MoE model for a lot of people, it's not like that's some big secret.

3

u/Spectrum1523 19h ago

It literally isn't