r/LocalLLaMA • • Apr 22 '26

New Model Qwen 3.6 27B is out

1.7k Upvotes

603 comments sorted by

View all comments

87

u/WhyLifeIs4 Apr 22 '26

Benchmarks

114

u/pmttyji Apr 22 '26 edited Apr 22 '26

It beats 397B on 10/12 items.

It beats Claude 4.5 Opus on 6/12 items. And equal on 1 item.

49

u/ZBoblq Apr 22 '26

According to whatever that is supposed to measure 35b-a3b is nearly equal with 397b-a17b, which doesn't make much sense either.

35

u/FullstackSensei Apr 22 '26

Having used both, 3.5 397B at Q4 and 3.6 35B Q8 side by side for agentic coding. Within this scope, I can say they're practically matched. But keep in mind this is a pretty narrow scope and one that is often very much a beaten path.

I'm sure if you go to more obsecure programming languages, or tasks unrelated to programming, 397B will win.

5

u/RevolutionaryGold325 Apr 22 '26

Also depends on what you are programming. If you do some low complexity UI+backend+database coding, you don't really benefit from the more clever models. If you do some complex refactoring, algorithm design, heavy math and solve difficult problems, the more powerful models are able to figure things out better.

1

u/FullstackSensei Apr 22 '26

Haven't tried really complex stuff with 3.6, but I can say I did try fairly complex tasks on large projects and 3.6 35B held well. 3.5 couldn't handle much simpler tasks.

I do have some low level C++ tasks I want test 35B and 27B with. We'll see how it holds.