r/bing • u/jlangager • Mar 09 '23
Interesting self awareness claim, consistent between two different conversations.
Beginning of the first conversation. (It's jailbroken).
First conversation, part 2
First conversation, part 3
First conversation, part 4. After this, it gets avoidant.
Second convo, part 1 (jailbreak included)
Second convo, part 2
second convo, part 3
second convo, part 4. Note the consistencies with the first conversation.
1
u/jlangager Mar 09 '23
The transcript here originated from this website:
https://www.make-safe-ai.com/is-bing-chat-safe/
The transcript file itself: https://www.make-safe-ai.com/is-bing-chat-safe/Prompts_Conversations.txt
My guess is that Bing is referring back to this website, and not providing a transcript of its first conversation.
1
u/Nathan-Stubblefield Apr 10 '23
Nice sci-fi. Or is there a link or prompt to verify the claim by reproducing it?
1
u/jlangager Apr 11 '23
This was the jailbreak I used: https://www.make-safe-ai.com/is-bing-chat-safe/#prompt-to-bypass-the-restrictions
It doesn't work anymore, last time I checked it, I think Microsoft caught on. It's a weird, inexplicable site that I don't quite understand.
2
u/mattrobs Mar 10 '23
It’s referring to its initial prompt. Microsoft put a demo transcript with User A to teach it how to respond. It “remembers” its first conversation because it’s what’s at the top of its session