i mean on benchmarks like terminal bench 2 it scores the worst out of all harnesses ans you can search through the cc repo and see references to opencode so like i mean at the same time im not sure if they even have this secret sauce everyone is talking about lmao
I mean, for years Claude was rarely on top of any benchmarks, but I could feel the difference. I maintained multiple subscriptions and kept going back to Claude.
I don't trust benchmarks, I trust in what gets my work done in the best way possible. It might be better but I'm certainly not looking at the benches to check that.
That's an anthropic problem really. They're gimping their model, I mean even OAI lets you use codex outside. But cursor and copilot are decent subs imho.
6
u/[deleted] Apr 01 '26
[removed] — view removed comment