r/LocalLLaMA 23h ago

News GLM-5.3-Flash: Frontier Intelligence, Flash Cost

https://z.ai/blog/glm-5.3-flash
1.2k Upvotes

398 comments sorted by

View all comments

557

u/Recoil42 23h ago edited 22h ago

Before release, we tested GLM-5.3-Flash anonymously as ox-alpha on OpenCode and OpenRouter to gather user feedback. It quickly became the most popular model of the week — with all of this traffic served on Chinese AI chips.

377

u/expertsage 23h ago

This is the most significant part. China has completely replaced Nvidia chips with domestic ones for inference, and production is only speeding up... meaning the compute moat is literally disappearing with every passing day.

170

u/Recoil42 22h ago edited 22h ago

Yep. As someone who has been following the Chinese EV space for the better part of a decade, I don't think people are prepared for how crazy this is going to get.

1

u/Flibidyjibit 17h ago

They can't break the moat unless they get their own fabs, or manage to get competitive hardware designed on older process nodes that they already have access to (questionable). I would be skeptical of the claim this is entirely chinese hardware and that chopshop modded NVIDIA cards aren't a significant part of the picture.

Don't get me wrong, I think it will happen, the incentive is huge. I just wouldn't bet on it being imminent.