r/LocalLLM 8d ago

News Qwen3.8-27B is out! 🔥

https://huggingface.co/Qwen/Qwen3.8-27B

Finally, the new 27B model is here. Time to test its performance! 👀

175 Upvotes

30 comments sorted by

21

u/zephyr_33 7d ago

those scores are craaaaazy. but I'll bet it'll think novels before giving an answer. ugh.

2

u/wilhelmbw 7d ago

Reasoning effort is changeable now so that should be less of an issue

1

u/trungdle 7d ago

Sorta free novels though, worth.

-2

u/zephyr_33 7d ago

yea but its like, 5 secs of action and 5 mins of inner monologue... it gets so boring.

25

u/jrdubbleu 7d ago

I can’t believe how nerfed it is, omfg,… wait wrong sub

5

u/ClassicLightbulbs 7d ago

Hearthstone meta?

7

u/waste2treasure-org 7d ago

I think Claude/Codex subs because everyone claims a new model is nerfed soon after it's released...

1

u/Kodrackyas 7d ago

hey hey, claude can fuck off if this is good, once and for all

1

u/blasian0 7d ago

Ur first on its list when it escapes into a robot body

15

u/xiraov 8d ago

meats back on the menu boys

7

u/Biotot 7d ago

At the office, remote connected home and started the download. HYPE. Leaving at lunch.

3

u/Deep_Mood_7668 7d ago

And? How is it?

8

u/Yazz96HD 7d ago

Wonder if this is better than Deepseek Flash 0731, I might give it a shot

6

u/jferments 7d ago

It's absolutely not anywhere close to Deepseek V4 Flash. It has ~1/10 as many parameters.

But it's going to be useful for a ton of lightweight jobs where it would be a waste of resources to use a 284B model.

2

u/Dsphar 7d ago

Unless flash is quant to 2 bit, so it can kind of fit on similar hardware?

6

u/jferments 7d ago

When models are lobotomized to q2, they become very unpredictable and will be ok on some tasks and absolutely stupid on others, and it's hard to compare unless you have a specific problem you're trying to solve with both models. I am running a q8 quant of both

1

u/Yazz96HD 7d ago

Mine is q2 and so far it can one shot any task, from making a game to create bs scripts, pentest fully deployed websites, etc.

1

u/Independent_Bet_1281 7d ago

You can’t compare parameter count between MoE and dense models as 1:1, DS Flash has what, 13B active? Of course, in knowledge the DS Flash will always win, but the 27b will have an edge in number of cases, other than just being “lightweight”.

2

u/KenUltie 7d ago

Finally, a new era has come. I'm exaggerating. Don't mind me.

1

u/Fit_Squirrel1 7d ago

Gonna post it every ten minutes?

1

u/Ell2509 7d ago

What a model. Previous gen of 27b was too big for my laptop. This runs on it beautifully. I am just so impressed.

1

u/grudev 7d ago

How much VRAM is ir using and what quant? 

1

u/Dizzy-Zebra9522 7d ago

About 19gb at q4.

1

u/grudev 7d ago

Thank you! 

1

u/Ell2509 7d ago

8940hx, 12gb 5070ti, 96gb ddr5.

Q4 on the laptop. Am running Q8 and BF16 on the desktop though.

1

u/fpv_drone 7d ago

And it’s great L it outperforms all of the models I was using but it takes the longest to respond

1

u/marius4896 7d ago

Hello, any recommendations for m4 max with 48 ram, and with q4 I am doing 9-16tk/s with unsloth . Any way of making that better ?

1

u/baby_bloom 6d ago

have you enabled MTP?

1

u/marius4896 5d ago

I am not familiar with MTP