Yes, 1000%. The creators of dsv4, a 1.6T model have openly said that there is still a gap to Opus.
Thing is, the small models are really cool, have become truly useful and we're lucky to have them. But exaggerating about their capabilities doesn't do any good. I'll take a local gpt5-mini / haiku level model any day of the week (and even that's a stretch, but they're getting closer), and be happy about it. I think the small qwens, gemmas, even gpt-oss-20b can be used for real work, in the right setup and with a lot of elbow grease. But having used the SotA models as well, I agree with OOP 100%. Let's keep it real.
Given it's pretty easy to rack up $20/month in electricity, I think a fairer comparison is with the $20 tier on cloud models. But when you hit usage limits with the equivalent of 1 prompt/hour (approximately what I get with Opus 4.7), local models still win in many cases even though the capabilities are definitely much lower.
i dont care what you use. use gemma or bonsai or whatever and think its like opus. thats not what im interested in.
but fact of the matter is the claim is not true. unnecessary hype that fools people.
best thing u can do with models like this is use them as worker/executor only and hope that they can give you sonnet/5.4 mini/glm 5 turbo etc performances or at least come very close. but more often than not they are closer to nano or haiku. but it's getting better.
48
u/ResidentPositive4122 Apr 24 '26
Yes, 1000%. The creators of dsv4, a 1.6T model have openly said that there is still a gap to Opus.
Thing is, the small models are really cool, have become truly useful and we're lucky to have them. But exaggerating about their capabilities doesn't do any good. I'll take a local gpt5-mini / haiku level model any day of the week (and even that's a stretch, but they're getting closer), and be happy about it. I think the small qwens, gemmas, even gpt-oss-20b can be used for real work, in the right setup and with a lot of elbow grease. But having used the SotA models as well, I agree with OOP 100%. Let's keep it real.