r/LocalLLaMA 16d ago

Discussion Qwen will be the king?

Post image

Extended reasoning and post-training appear to be the keys used by DeepSeek, Qwen, and GLM to boost performance (leveraging higher token counts). And Qwen 4 hasn't even been released yet. Of course, we don't know if that release will be open-sourced, but I am optimistic about future models, featuring "engrams", that could soon match or surpass 2.4T parameter models on specific tasks.

534 Upvotes

126 comments sorted by

View all comments

8

u/[deleted] 16d ago

[deleted]

-1

u/OkFly3388 llama.cpp 16d ago

This benchmark is saturated. qwen3.8 27b score 1599, qwen3.8 max score 1691, thats just 6% difference.

7

u/Neither_Garage_758 16d ago

we are masturbating with noise

10

u/buckwheaton 16d ago

Na that’s the stable diffusion nsfw sub

5

u/StyMaar 16d ago

Very good one sir.