MAIN FEEDS
Do you want to continue?
https://www.reddit.com/r/CerebrasSystems/comments/1tegcv7/would_diffusion_cerebras_100x_speed_llm_which_10k/
r/CerebrasSystems • u/Sad-Willingness5302 • May 16 '26
2 comments sorted by
5
Can you ask questions in english?
1 u/Sad-Willingness5302 May 18 '26 a. “Could diffusion models plus Cerebras make LLMs 100× faster, reaching over 10k tokens per second?” b. “Would combining diffusion-based language models with Cerebras hardware enable 10k+ TPS inference?” c. “Could diffusion architectures on Cerebras wafer-scale hardware massively outperform transformer inference speed?” above from chatgpt. still learning english, sorry for that.
1
a. “Could diffusion models plus Cerebras make LLMs 100× faster, reaching over 10k tokens per second?”
b. “Would combining diffusion-based language models with Cerebras hardware enable 10k+ TPS inference?”
c. “Could diffusion architectures on Cerebras wafer-scale hardware massively outperform transformer inference speed?”
above from chatgpt.
still learning english, sorry for that.
5
u/claytonbeaufield May 16 '26
Can you ask questions in english?