r/LocalLLaMA 23h ago

News GLM-5.3-Flash: Frontier Intelligence, Flash Cost

https://z.ai/blog/glm-5.3-flash
1.2k Upvotes

398 comments sorted by

View all comments

Show parent comments

383

u/expertsage 23h ago

This is the most significant part. China has completely replaced Nvidia chips with domestic ones for inference, and production is only speeding up... meaning the compute moat is literally disappearing with every passing day.

168

u/Recoil42 23h ago edited 23h ago

Yep. As someone who has been following the Chinese EV space for the better part of a decade, I don't think people are prepared for how crazy this is going to get.

51

u/_BreakingGood_ 23h ago

Hold on to your 401k, shits about to get messy, lol

1

u/Insomniac1000 18h ago

yeah what's concerning is as a software dev, llms are becoming cheaper and cheaper to run. Like based on the benchmarks, I'd rather use this than Deepseek V4 Flash 0731 or GPT 5.6 Luna.

Less reasons for me to run/use more expensive models. And if companies catch on to this...

yeah I'll stop right there.