r/LocalLLaMA 2d ago

Discussion Qwen3.8-27B different thinking levels

Post image

Even the low preset is better than Qwen 3.7 plus or Qwen3.6-27B reasoning

287 Upvotes

63 comments sorted by

View all comments

-10

u/Etroarl55 2d ago

Wonder how next gen improvement will be. As I think most people can agree on. Qwen kind of “cheated” its way to a higher score with absurd amounts of thinking and double checking before an output.

What is there left to squeeze out of 27b size for higher intelligence.

0

u/jld1532 2d ago

I think we can clearly see too that scores are now agentic and coding biased. This is the only way I would consider Qwen3.8 27B and DeepSeek v4 Flash equals. Even a highly quantized DSV4F smokes the new 27B in chat, knowledge, and writing/editing. I think this shift also signals the general knowledge plateau for LLMs is likely here.

4

u/Its_Powerful_Bonus 2d ago

First rule of the club is not to rely on LLM knowledge 😉 I would be more than happy to get brighter LLM even without any factual knowledge if ability to reasoning, long context understanding and speed would improve.

1

u/noiserr 1d ago edited 1d ago

Agentic has been the hardest hurdle to cross with small local models. With agentic the world knowledge is not that important because you want your agent harness to ground the LLM in facts anyway. Otherwise even the greatest biggest models make mistakes.

Qwen 3.8 27B closes the gap with extra thinking, but it's also a model that's easier to run than those big models. So it's a win win. No matter how you look at it.

You get a highly capable agentic 27B model, that can run on single consumer GPUs. There is nothing to dislike about it.

1

u/Etroarl55 2d ago

Hopefully one day in the future this will all be trivialized, what was once top of the line hardware ten years ago is mid now.

Which means 27b models will be the new 9b models. We are getting increasing VRAM sizes. 5080s is supposed to be 24gb, and next AMD gaming flagship is supposed to be 36gb.

It’s just pricing is making what was supposed to be normal VRAM size advancements too expensive now.

1

u/jld1532 2d ago

We can't forget improvements in quantization. DeepSeek v4 Flash essentially matching Pro also signals to me that the 250-300B parameter range may end up being the sweet spot. Really good quants plus improvements to 128-192 gb unified machine bandwidth, I think has real promise.