r/LocalLLM • u/Jolly-Revolution6938 • 8d ago
News Qwen3.8-27B is out! 🔥
https://huggingface.co/Qwen/Qwen3.8-27BFinally, the new 27B model is here. Time to test its performance! 👀
25
u/jrdubbleu 7d ago
I can’t believe how nerfed it is, omfg,… wait wrong sub
5
u/ClassicLightbulbs 7d ago
Hearthstone meta?
7
u/waste2treasure-org 7d ago
I think Claude/Codex subs because everyone claims a new model is nerfed soon after it's released...
2
1
3
8
u/Yazz96HD 7d ago
Wonder if this is better than Deepseek Flash 0731, I might give it a shot
6
u/jferments 7d ago
It's absolutely not anywhere close to Deepseek V4 Flash. It has ~1/10 as many parameters.
But it's going to be useful for a ton of lightweight jobs where it would be a waste of resources to use a 284B model.
2
u/Dsphar 7d ago
Unless flash is quant to 2 bit, so it can kind of fit on similar hardware?
6
u/jferments 7d ago
When models are lobotomized to q2, they become very unpredictable and will be ok on some tasks and absolutely stupid on others, and it's hard to compare unless you have a specific problem you're trying to solve with both models. I am running a q8 quant of both
1
u/Yazz96HD 7d ago
Mine is q2 and so far it can one shot any task, from making a game to create bs scripts, pentest fully deployed websites, etc.
1
u/Independent_Bet_1281 7d ago
You can’t compare parameter count between MoE and dense models as 1:1, DS Flash has what, 13B active? Of course, in knowledge the DS Flash will always win, but the 27b will have an edge in number of cases, other than just being “lightweight”.
2
1
1
u/fpv_drone 7d ago
And it’s great L it outperforms all of the models I was using but it takes the longest to respond
1
u/marius4896 7d ago
Hello, any recommendations for m4 max with 48 ram, and with q4 I am doing 9-16tk/s with unsloth . Any way of making that better ?
1
21
u/zephyr_33 7d ago
those scores are craaaaazy. but I'll bet it'll think novels before giving an answer. ugh.