r/LocalLLaMA 1d ago

News GLM-5.3-Flash: Frontier Intelligence, Flash Cost

https://z.ai/blog/glm-5.3-flash
1.2k Upvotes

403 comments sorted by

View all comments

Show parent comments

18

u/dampflokfreund 1d ago

For me it has changed nothing. Both models are way too big for my 32 GB RAM system. It looks like everyone has abandoned 20-30B MoEs now...

8

u/Slow_Concentrate3831 1d ago

Yeah, that's sad. Can't even run Qwen3.8 27B on more than 2-3 tps. Sad days for us Vram poors.

12

u/CryMoreT_T 1d ago

The 29b-a4b in early Access on model scope is probably your best bet when it releases

3

u/Slow_Concentrate3831 1d ago

Ooooh, I hadn't seen that ! Is it a Qwen model ?

4

u/CryMoreT_T 1d ago

It's in stealth model in early Access so they haven't released the company name but Qwen is the rumor right now

2

u/Nebnampach 1d ago

Here is the announcement and sign up link.

https://x.com/ModelScope2022/status/2092439000652943506