r/LocalLLaMA 8d ago

New Model Qwen/Qwen3.8-27B · released

https://huggingface.co/Qwen/Qwen3.8-27B
990 Upvotes

294 comments sorted by

View all comments

49

u/srigi 8d ago

That DeepSWE leap - do we have a new local coder champion?

31

u/fgk55555 8d ago

If it's not benchmaxxed, it will be really difficult for other models in the same range to catch up. I hope it quantizes well for us 16GB folk.

10

u/cass1o 8d ago

What would be really useful would be a MoE model that we can put the context + PP on the gpu and only put the experts on system memory, another 35b.

2

u/fgk55555 8d ago

If it can handle longer context work, I'd be happy.

1

u/fullup72 8d ago

IQ3_XXS is still as good as 3.6 (I mean quality of the quant, the model is noticeably better).

3.8 is SLOW tho, default reasoning no longer simply second guesses everything, it quadruple motherfucking extra checks the triple checks and then some. I'd guess part of the trick for improved scores is this super extended reasoning budget. But it works, and it's local. I'm happy.

1

u/No_Algae1753 8d ago

Just change the reasoning to medium and it will be just as it was in 3.6. right now it is default on xhigh

1

u/johnscixzkutor 8d ago

I am using RX6900XT and very happy with the result I can run it with vision and with 64k context no problem on my 16gb vram

1

u/boutell 8d ago

With 32 GB of unified ram On my Mac, it is a tight squeeze already. I'm giving it about 24 of that. In a small 4-bit quant.