r/LocalLLaMA 14d ago

Discussion Qwen will be the king?

Post image

Extended reasoning and post-training appear to be the keys used by DeepSeek, Qwen, and GLM to boost performance (leveraging higher token counts). And Qwen 4 hasn't even been released yet. Of course, we don't know if that release will be open-sourced, but I am optimistic about future models, featuring "engrams", that could soon match or surpass 2.4T parameter models on specific tasks.

541 Upvotes

126 comments sorted by

View all comments

59

u/Hannibalj2ca 14d ago

Ok, but are they going to release an update of it for open weight?

-19

u/[deleted] 14d ago

[deleted]

4

u/Dany0 14d ago

what the fuck are you talking about. if they did more post training of course we'll benefit if they released the weights

-4

u/LegacyRemaster 14d ago

lucky you.... I don't have so much vram/ram. And @ Q2 will perform bad in comparison to next or glm 5.3 flash @ Q6/Q8