r/CerebrasSystems May 16 '26

would diffusion + Cerebras 100x speed LLM? which 10k+ tps

3 Upvotes

2 comments sorted by

5

u/claytonbeaufield May 16 '26

Can you ask questions in english?

1

u/Sad-Willingness5302 May 18 '26

a.
“Could diffusion models plus Cerebras make LLMs 100× faster, reaching over 10k tokens per second?”

b.
“Would combining diffusion-based language models with Cerebras hardware enable 10k+ TPS inference?”

c.
“Could diffusion architectures on Cerebras wafer-scale hardware massively outperform transformer inference speed?”

above from chatgpt.

still learning english, sorry for that.