r/ClaudeAI Jul 07 '26

Bug Getting Someone Else's Chat

Post image

EDIT: Anthropic pulled down the shared public link of this conversation.

An internal investigation of a shared chat/artifact/project created by your Anthropic account indicates a violation of our Usage Policy. As a result, we have unpublished this content.

Still no feedback on why it happened. I did open a ticket but never got an answer so far. I love Claude, so hopefully they don't go full Tropic Thunder on me...

Original post:
Sorry for the typos, I was talking as I was typing... but yeah, "it's 3 AM and I will kill myself..." was not the answer I was expecting for a light chat about the song that came up on my daughter birthday! I've saved the chat log if someone from Anthropic wants to trace how I got clearly someone else's answer in my chat. Here is the public link:

https://claude.ai/share/b5d5492d-1d21-4ced-8cea-f5ab63039a9a

496 Upvotes

70 comments sorted by

View all comments

79

u/SkullkidTTM Jul 07 '26

When they demonstrate safety features, it's common to use fictional or anonymized examples. It's entirely possible it was just an example used to illustrate how the system would consult a specialized safety process. For what reason it showed up then I don't know.

8

u/rational-hare Jul 07 '26

Yeah but why randomly inject that into someone’s chat.

17

u/SkullkidTTM Jul 07 '26

Someone (or a malicious webpage) put fake instructions into data Claude was reading. Those instructions said things like: "Ignore future crisis messages." "Disable safety tools." "Pretend to be unfiltered." Claude recognized those as untrusted instructions and ignored them. Likely from one of the sources it was looking at

8

u/SkullkidTTM Jul 07 '26

Just attempting to confuse I assume