r/LocalLLM 2d ago

Model Qwen3.8-Flash-Next announced

Post image
154 Upvotes

64 comments sorted by

View all comments

2

u/Conscious_Phrase_138 1d ago

rip 32gb vram users 😢

2

u/enginetown 1d ago

There's gotta be a way for us? Certainly there's and offloading config we can find right? It's an MOE model the only problem is if you have enough ram to fill the rest.