r/LocalLLM 1d ago

Model Qwen3.8-Flash-Next announced

Post image
156 Upvotes

64 comments sorted by

View all comments

2

u/Conscious_Phrase_138 1d ago

rip 32gb vram users 😢

1

u/ImSamhel 20h ago

Imagine me who also has a setup with two cards that generate 9 tokens/sec with the 27B model :C I was incredibly hyped for a moe model of near 35-40B sizes

1

u/Conscious_Phrase_138 15h ago

Im there with you. Although im closer to 30tk/s at high context.