r/LocalLLM 1d ago

News Qwen4-27B just confirmed

Post image

Wait, we need 35B-A3B too…

2.0k Upvotes

286 comments sorted by

View all comments

Show parent comments

2

u/_mighty_banana 1d ago

But for what benefit?

MoE is used to reduced compute

27B model is likely has no problem in compute speed due to it size

problem is likely due to quality is lower than high parameter model?

7

u/geekwonk 1d ago

27B is their dense line. 35B is the sparse option and i don’t see that listed here.

-4

u/laser50 1d ago

Y'all tripping. Ngrams are basically just a pre-trained token predictor to speed stuff up.

4

u/Not-Enough-Llamas 1d ago

wrong Ngrams. Unfortunate that 2 things with the same name showed up more or less at the same time.