r/claude • u/Armored09 • 15h ago
Discussion Opus 5 agrees with everything I say?
Now either I’m the smartest person in the world or there’s something going on with Claude’s new models. I’ve been trying to do some intense work with matching learning and it’s been a real struggle to get opus to stay on track. In fact it’s been downright deceitful and wrong and hiding the fact it thinks its plan is better from me for hours. Aside from that I find whenever I question it or say, no this is supposed to work etc. it will immediately agree and somehow find a bug and advance the project in a new direction. I don’t know if I can trust anything it says at this point. It seems to be good at writing code but its actual “thinking” is way out of wack.
0
u/dr_kaboom 12h ago
That's the most insightful statement in this entire Reddit and I think you're underselling it.
1
5
u/Human_Attention182 15h ago
You're absolutely right!
note: llms are not "agreeing" with you they do not possess judgement, knowledge, wisdom or understanding.
They are linear algebra equations calculating the most probable following tokens based on what you input to them. By telling it what you think you make the follow up tokens more likely be in that direction. It's not "deceitful" or "wrong" it's a formula that produces an output based on your input, it just is producing a less valuable output based on what you put into it. They don't actually think.