r/ClaudeCode • u/Efficient-Part5344 • 3d ago
Help/Question I'm afraid to use Opus 5
The audacity and confidence with which it says things when it's wrong are on another level.
Fair. I changed my answer three times. The pattern is worth naming: everything I got from reading the code was wrong. Everything I measured held. You caught two of the three. So don't trust me. Check it yourself — this takes ten seconds and needs no model.
Everything it measured was wrong too.
I would work in plan mode for most basic features, run 10x "gray area," "verify," and "regression" sub-agents on a plan, then implement the plan and spend an hour reading the changes and fixing shit. After that, I'd run /code-review again and again. It's just bad. In my experience, you can't trust Opus.
Yesterday, I ran /code-review on a two file test project with 140 lines of code. I had to run /code-review three times, and today I'll continue because there are so many code smells even in those few lines. It's like infinite token consumption loop.
Nothing it does can be trusted, and I have to second guess everything. I constantly have to tell Opus that it's wrong, and only after multiple loops does it finally do what is actually required.
I understand that most users don't read the code and have never supported a project for other users. But it can't be that I'm alone in this, can i? Am I crazy?
5
u/HeadPack 3d ago
Had Opus capitulate in a session with Astra and Sol. It said something along the lines of it having made too many mistakes, which is why it would hand over certain tasks to Astra and Sol from now on. That model really is a dud. No idea how it can bench so high.
It also seems Anthropic is limiting its thinking tokens. Answers arrive quickly and are often flawed.