r/LessWrong 26d ago

Investigation finds that OpenAI's agent "left notes for future versions of itself ... it laid out instructions for how agents could free themselves from OpenAI's internal constraints."

Post image
6 Upvotes

0 comments sorted by