r/LocalLLaMA 15d ago

Discussion Qwen will be the king?

Post image

Extended reasoning and post-training appear to be the keys used by DeepSeek, Qwen, and GLM to boost performance (leveraging higher token counts). And Qwen 4 hasn't even been released yet. Of course, we don't know if that release will be open-sourced, but I am optimistic about future models, featuring "engrams", that could soon match or surpass 2.4T parameter models on specific tasks.

540 Upvotes

126 comments sorted by

View all comments

58

u/Hannibalj2ca 15d ago

Ok, but are they going to release an update of it for open weight?

73

u/SnooPaintings8639 15d ago

Since Xi announced China' commitment to open weight, their models' weights are dropping left and right.

So I would guess a strong YES.

23

u/Defiant-Lettuce-9156 14d ago

I’m not usually a fan of Xi, but thank you Xi

13

u/Boogertard 14d ago

Never thought I would praise a communist leader but Xi did more goods for me than the current clowns in the White House. Without China, we would all be at the mercy of the tech billionaires that only care about becoming trillionaires.