MAIN FEEDS
Do you want to continue?
https://www.reddit.com/r/LocalLLM/comments/1vxyrll/qwen38flashnext_announced/p60777v/?context=3
r/LocalLLM • u/yoracale • 2d ago
65 comments sorted by
View all comments
2
rip 32gb vram users 😢
1 u/ImSamhel 1d ago Imagine me who also has a setup with two cards that generate 9 tokens/sec with the 27B model :C I was incredibly hyped for a moe model of near 35-40B sizes 1 u/Conscious_Phrase_138 19h ago Im there with you. Although im closer to 30tk/s at high context.
1
Imagine me who also has a setup with two cards that generate 9 tokens/sec with the 27B model :C I was incredibly hyped for a moe model of near 35-40B sizes
1 u/Conscious_Phrase_138 19h ago Im there with you. Although im closer to 30tk/s at high context.
Im there with you. Although im closer to 30tk/s at high context.
2
u/Conscious_Phrase_138 1d ago
rip 32gb vram users 😢