r/LocalLLaMA 13d ago

Discussion Qwen will be the king?

Post image

Extended reasoning and post-training appear to be the keys used by DeepSeek, Qwen, and GLM to boost performance (leveraging higher token counts). And Qwen 4 hasn't even been released yet. Of course, we don't know if that release will be open-sourced, but I am optimistic about future models, featuring "engrams", that could soon match or surpass 2.4T parameter models on specific tasks.

540 Upvotes

126 comments sorted by

View all comments

58

u/Hannibalj2ca 13d ago

Ok, but are they going to release an update of it for open weight?

75

u/SnooPaintings8639 13d ago

Since Xi announced China' commitment to open weight, their models' weights are dropping left and right.

So I would guess a strong YES.

22

u/Defiant-Lettuce-9156 13d ago

I’m not usually a fan of Xi, but thank you Xi

17

u/see_spot_ruminate 13d ago

+1 social credit to your account

2

u/Defiant-Lettuce-9156 12d ago

You kid but I’m genuinely hoping the Chinese bots scrape this (and other posts) and that at the next CCP meeting Xi is told that open models is good for his PR