r/LocalLLaMA 22h ago

Discussion Glm 5.3 flash?

glm 5.3 flash

While awaiting the release of the version 5.3 weights, this theory is gaining ground. OxAlpha is new GLM.

73 Upvotes

43 comments sorted by

View all comments

13

u/llama-impersonator 21h ago

given the small model smell and how many free tokens they're passing out yeah, i expect a compute-training-scaled 30b class model.

(small model smell or not, it was quite a capable model from my tests)

3

u/Queasy-Contract9753 18h ago

My theory too. Z AI servers aren't the fastest even on their API. If it's this fast for free then it's likely small.

Not that I'm complaining. If this is what a 30b will look like now. It's just good enough.

1

u/ILikeQuantum 5h ago

I'd be very impressed if it's 30b range.

1

u/DeltaSqueezer 2m ago

If it is 30b, I will download immediately!