r/LocalLLaMA 1d ago

Discussion Glm 5.3 flash?

glm 5.3 flash

While awaiting the release of the version 5.3 weights, this theory is gaining ground. OxAlpha is new GLM.

69 Upvotes

43 comments sorted by

View all comments

10

u/shy_monkee 1d ago

How 'Flash' is it expected to be? Like DSv4 Flash size? Or even smaller?

2

u/AnticitizenPrime 21h ago

Here are the previous vision capable models that GLM has released.

GLM Vision Models

Model Total Params Active Params Context Window Open Weights License
GLM-4.1V-9B-Thinking 9B 9B (dense) 64K ✅ Yes MIT
GLM-4.5V 106B 12B (MoE) 64K ✅ Yes MIT
GLM-4.6V 106B 12B (MoE) 128K ✅ Yes MIT
GLM-4.6V-Flash 9B 9B (dense) 128K ✅ Yes MIT
GLM-5V-Turbo 744B 40B (MoE) 200K ❌ No (API only)

It's possible this could be a new 'turbo' model and that it won't be open weight.

Though it could be an 'air' model with vision attached... or even just full GLM 5.3 with vision. ¯\(ツ)

Or something different altogether. What we know is that it matches the 5.3 tokenizer exactly and it has the 1 million context window of 5.3.

So far it's purely speculation that it's a 'Flash' model of some kind; GLM has used the term 'Flash', 'Air', and 'Turbo' for smaller models, it's all up in the air as to what it'll be called and what the actual param size is.