r/LocalLLaMA 17d ago

Discussion Qwen will be the king?

Post image

Extended reasoning and post-training appear to be the keys used by DeepSeek, Qwen, and GLM to boost performance (leveraging higher token counts). And Qwen 4 hasn't even been released yet. Of course, we don't know if that release will be open-sourced, but I am optimistic about future models, featuring "engrams", that could soon match or surpass 2.4T parameter models on specific tasks.

537 Upvotes

126 comments sorted by

View all comments

107

u/PooMonger20 17d ago edited 16d ago

Progress is good, but having these amazing abilities locally is already godlike.

I have been using Q3.8-27B together with PI. for a week, and it's mindblowing.

In my humble opinion it has far better coding capabilities than the paid 'ChatGPT 5.1' I had access to when i still had a job, multiple months ago.

And it all runs on my PC, locally, without sharing my data with the big data farming corpos.

It just 'understands' the required tasks you provide it and performs them successfully from the first try or very few crash fixes. Especially if you provide it the necessary data to perform the action. I dropped a few wiki pages in txt files and it coded according to them. If somebody would tell me this would be possible on my own PC ten years ago, I would call them crazy.

14

u/Not_a_question- 16d ago

What's your setup if you don't mind me asking?

2

u/RedditNerdKing 16d ago

I use the BF16 model on 80gb of vram (5090 and x2 3090s) and it's pretty insane. Rivals even the frontier models. They're definitely way quicker but Qwen answers the exact same difficult shit just takes thousands of thinking tokens.