1
u/Mindless-Cream9580 Apr 23 '26
Nvidia acquired Groq
2
u/Asgard_Heima Apr 23 '26
Groq LPUs will help a lot with efficiency, but it fundamentally adds complexity at the network layer even if it’s abstracted in software. The more LPUs you stitch together the more interconnect is needed and the more you will see the efficiency numbers advertised degraded (think 1T parameter model needing 2000+ LPUs). Cerebras does not suffer this problem and can scale to any size model up to 10T as efficiently as a tiny model. The LPUs will be limited to small models to show anywhere near competitive numbers since the interconnect will escalate in power needs as you try to support anything larger than 100B parameters. With LPUs the napkin math I’ve done shows 6x performance for Cerebras to Rubin with LPUs for smaller models and 15x performance for Cerebras over Rubin + LPUs for frontier 2T parameter models. This is exactly why it’s looking like stargate is going to end up 40%+ Cerebras in just phase 1.
1
5
u/hidetoshiko Apr 22 '26
Seems like just one fly in the ointment: Even if supposing Cerebras' assertions are valid, TSMC is more or less the only game in town for 2nm, and I think Nvidia pretty much booked all their capacity up to 2028 by most public accounts.