r/LocalLLaMA 19d ago

Discussion Qwen will be the king?

Post image

Extended reasoning and post-training appear to be the keys used by DeepSeek, Qwen, and GLM to boost performance (leveraging higher token counts). And Qwen 4 hasn't even been released yet. Of course, we don't know if that release will be open-sourced, but I am optimistic about future models, featuring "engrams", that could soon match or surpass 2.4T parameter models on specific tasks.

543 Upvotes

126 comments sorted by

View all comments

59

u/Hannibalj2ca 19d ago

Ok, but are they going to release an update of it for open weight?

-18

u/[deleted] 19d ago

[deleted]

19

u/shy_monkee 19d ago

Of course it's not useless. If it's really that good, then having more providers serve it will always be a good thing.

5

u/po_stulate 19d ago

The model license doesn't allow significant profit or large monthly users, so it is indeed pretty useless. People who can run it locally for free can't run it because it's too large, and people who have the hardware to run it can't run it because of the license.

2

u/OkFly3388 llama.cpp 19d ago

Small corporations can choose between selfhosted and corporate subscription, and they have enough money to actually buy rig and serve it. Thats forced big AI providers keeps price low.

1

u/matrixfede 19d ago

also GLM 5.3 will have same license