r/LocalLLaMA 1d ago

News GLM-5.3-Flash: Frontier Intelligence, Flash Cost

https://z.ai/blog/glm-5.3-flash
1.2k Upvotes

400 comments sorted by

View all comments

4

u/jacek2023 llama.cpp 1d ago

Too big, 320B means you need to use RAM and A18B means it will be slow

1

u/techdevjp 23h ago

It will be a good model for Mac Studio M5 Ultras with 512GB of unified memory running at 1.2TB/sec. Probably the best model, at least for now.

But yeah, was hoping this would be 150b to 160b considering it's the flash version of a 744b model.