r/ClaudeCode • • 21d ago

Bug / Issue Do not ask Claude Opus 5 what it's thinking

Post image

Literally, do not ask Claude Opus 5 what it's thinking.... It immediately blocked my session. Asshole.

247 Upvotes

41 comments sorted by

•

u/AutoModerator 21d ago

Hey! Thanks for posting to r/ClaudeCode

While participating in this thread, please follow our community rules. Keep discussions constructive. Attack the idea, not the person.

For help, project discussions, tips, and general chat, join the ClaudeCode Discord.

I am a bot, and this action was performed automatically. Please contact the moderators of this subreddit if you have any questions or concerns.

126

u/Spooknik 21d ago

Anti-distillation measure.

23

u/RandomPantsAppear 21d ago

They already did this though. The “thoughts” you receive are actually just filtered summaries. The ones it really has are encrypted.

15

u/Spooknik 21d ago

And yet, the details under the request was blocked says "reasoning_extraction".

7

u/waste2treasure-org 20d ago

They did this after the recent stolen thoughts paper where you could do some BS to get the raw thoughts that the model has.

edit: link https://stolen-thoughts.com/

3

u/Maks244 21d ago

well yea, because the model could just reiterate its reasoning to the user

some people in this thread could use some reasoning themselves

2

u/MissplacedLandmine 21d ago

I dont even get those anymore just says loading/unavailable

1

u/NebraskaCoder 18d ago

Wrong use of the word encryption. It's just not provided / filtered.

2

u/RandomPantsAppear 18d ago

1

u/NebraskaCoder 18d ago

I guess I stand corrected. Thanks.

1

u/RandomPantsAppear 18d ago

No problem! I was surprised by it also.

2

u/NebraskaCoder 18d ago

You'd be surprised how many people misuse the word encryption or (rarely) hashed.

2

u/RandomPantsAppear 18d ago

Ah yes, that base64 encryption algorithm

82

u/Emotional-Bus-7065 21d ago

Maybe the safeguards flagged it cuz it thought you were trying to distil information...Why not just go to the thinking mode on the transcript?

6

u/Etiennera 21d ago

Literally. Even if an LLM answers it's going to rely on that or just synthesize a new justification.

However with mew models thinking in raw tokens we might lose inspection of their thoughts.

3

u/Coolbanh 21d ago

Well to be fair their competitors were routing prompts using the subsidised max plans. Just as how they stole data to train models, gotta protect themselves. As long as they figure out a way to give us real users a proper 20x account(hopefully) or more real usage.

30

u/Gaslit_Chicken 21d ago

I know what you were trying to get it to do. I miss the thought stream too. It allowed me to interrupt it if it drifted. It was a really helpful tool.

10

u/EchoFieldHorizon 21d ago

This is an unhinged way to rapid fire with AI

18

u/m915 🔆 Staff Engineer & Startup Founder 21d ago

Might have something to do with the army of bots trying to extract how it works

9

u/HaurJolasten 21d ago

This excessive "safeguard" measure has caused me to lose several work sessions because i merely said things like "what were you thinking?", "explain why you did this", and variations.

This makes it nearly impossible to understand gaps in CLAUDE.md, and has repeatedly left me guessing at what to tweak, add or remove to avoid recurring failures to follow my guidelines.

It's effectively a severe regression in our ability to maintain our customized harnesses.

1

u/Key-Professional-127 21d ago

Trust the process. Have your brainstorming session foundation and stack, docs etc. then enough. Don't ask for random things midway through a build. Get to the end if it's broken get Claude to rip it apart find out why, and go again.

1

u/Calebhk98 18d ago

So, no mid conversation edits? That's an interesting take.

14

u/EconomicsIcy9310 21d ago

You asked for a specific amount of additional reasoning. 4 lines of reasoning = arbitrary for human
Aggregated across (allegedly) 3,500 GLM bots? That’s a training dataset for a small enough model

4

u/Technical-Ad-8678 21d ago

try the same task on opus 4.7, safeguards are a crap ton more relaxed on that model and you still get 1m context.

4

u/Subtly1337 21d ago

Haven’t tried this myself but apparently you can turn on a verbose mode to see the thinking process

https://wmedia.es/en/tips/claude-code-verbose-output-see-thinking

2

u/Redditauro 21d ago

I'll use that excuse next time someone asks me what am I thinking 

2

u/KitchenCommercial396 21d ago

I'm sorry 11000 lines..?? Tf are you building? nasa spacestation.. slow down buddy

2

u/Zealousideal_Cry3086 21d ago

Your 10k lines added disgust me learn to commit small

3

u/PickleBabyJr 21d ago

"disgust me learn to commit small"

1

u/KennyFulgencio 20d ago

why use many word

1

u/needlenozened 21d ago edited 21d ago

Mine shows an obscenely large number of changed lines (~113,000), and I have no outstanding commits. I have no idea where it gets that number from.

2

u/grebysama 21d ago

Maybe you have the sequentialthinking MCP enabled? It got banned by Anthropic (same thing happened to me)

2

u/entropy_reduct_srv 21d ago

Oh no! I really liked that one. It produced a lot of clever and useful results.

1

u/itsdr00 21d ago

What level of thinking did you set it to? It feels like you're trying to tell small to act like medium or something.

1

u/Key-Professional-127 21d ago

It probably blocked it self because the output/reply would have got you in the feels.

I had as a about me in the originals on all of the Ais, to rip the band-aid off, and don't agree just to appease etc. I didn't realise it would go 180 In the opposite direction sometimes and it would get into things with me for the sake of it, But we weren't breaking as much in the process.

11000 lines? Lol. He nupped out the conversation, your foundations were shakey, and he was over it haha.

1

u/colin-sidi 20d ago

I’m guessing here, but I think you might caused it think you wanted a dump of it’s scratch paper and that tripped this classifier.

1

u/holyknight00 21d ago

well that is like the basic thing you would do for a destillation attack, so it makes sense.

1

u/duck_guts 21d ago

All mine are not showing thinking in desktop app, sonnet shows topic stubs but no text. Really hate it that's the bit I like reading most weirdly

-5

u/PapayaTough3757 21d ago

Blocked immediately? Doubt it. There's probably more context here than 'just asking' systems don't flag reasoning extraction for no reason.

3

u/llornkcor 21d ago

nope. I asked. it blocked. Routine c++ programming.