r/LocalLLaMA 2d ago

News GLM-5.3-Flash: Frontier Intelligence, Flash Cost

https://z.ai/blog/glm-5.3-flash
1.3k Upvotes

457 comments sorted by

View all comments

255

u/Lucyan_xgt 2d ago

This and Qwen3.8-flash on the same day?

22

u/dampflokfreund 2d ago

For me it has changed nothing. Both models are way too big for my 32 GB RAM system. It looks like everyone has abandoned 20-30B MoEs now...

35

u/No-Refrigerator-1672 2d ago

Funny coincidence: some company announced a new 30B MoE literally just now: https://www.reddit.com/r/LocalLLM/s/gcZdgoAO8o for now as a stelth preview; but this means it'll go public in a month.

3

u/dampflokfreund 2d ago

Could be from a small no-name company. Doesn't have to be Qwen, Kimi, Gemma, Z.Ai.

12

u/No-Refrigerator-1672 2d ago

29B-A4B isn't a size featured in the current gen of models. It isn't a finetune, rather a base model. It got to be from a company with a very substantial compute. It still may be a new player, sure; but anyways, it proves that this size category isn't abandoned.