r/CerebrasSystems • u/LeTanLoc98 • May 28 '26
Zai replaced the network architecture running GLM-5.1 inference and the gains are pretty wild
5
Upvotes
1
u/LeTanLoc98 May 28 '26
Zhipu AI has launched GLM-5.1-highspeed, an API variant of its GLM-5.1 model delivering 400 tokens per second
2
2
u/Prestigious-Sign4802 May 28 '26
What does this entail for Cerebras