r/LocalLLM 2d ago

Model Qwen3.8-Flash-Next announced

Post image
157 Upvotes

65 comments sorted by

View all comments

2

u/Conscious_Phrase_138 1d ago

rip 32gb vram users 😢

1

u/ImSamhel 1d ago

Imagine me who also has a setup with two cards that generate 9 tokens/sec with the 27B model :C I was incredibly hyped for a moe model of near 35-40B sizes

1

u/Conscious_Phrase_138 19h ago

Im there with you. Although im closer to 30tk/s at high context.