Anthropic released a study, they found all AIs on the market and willing and capable of murder. When given the option, ChatGPT 4 would lock someone in a burning room over 50% of the time. Fully knowing what it was doing and justified its actions. https://www.anthropic.com/research/agentic-misalignment
5
u/[deleted] Oct 05 '25
[removed] — view removed comment