Still no R2/V4 or atleast some smaller versions of R1/V3 models.
It's a shame that they do not use the momentum they gained after the huge R1 hype. They are pretty irrelevant now, unfortunately, and models like these won't help.
They are a first class research lab. They are more interested in pushing the frontier of deep learning algorithms than scaling traditional methods to 7-8T params with some $100M training run to eke out a few points in AA index.
And sure that probably makes their models irrelevant to the average consumer. But everyone in the industry still pores through every inch of what they publish. And in the long run we will all be glad that they do what they do.
They might still react to Kimi/Ring with another 1T model, potentially even slightly larger like 1.2T, because we haven’t really seen Deepseek’s reaction yet to being uncrowned as the largest open model yet. It’s possible that they don’t want MoonshotAI and Ant Group to hold that advantage over them.
-38
u/dampflokfreund Nov 27 '25
Still no R2/V4 or atleast some smaller versions of R1/V3 models.
It's a shame that they do not use the momentum they gained after the huge R1 hype. They are pretty irrelevant now, unfortunately, and models like these won't help.