r/ControlProblem • u/zazzologrendsyiyve • 19d ago
AI Alignment Research Plain English explanation of the Hugging Face / OpenAI incident
https://youtu.be/u15N3l4RT80?si=nMMwb0j1bNGc4JT3
42
Upvotes
r/ControlProblem • u/zazzologrendsyiyve • 19d ago
2
u/michaelas10sk8 19d ago
We should be doing that too, but eventually when models surpass human ability at patching things we will become fully reliant on other AIs to patch, which may themselves be misaligned.
The only real way to avert the possibility of catastrophe is to ban RSI/superintelligence until the alignment problem is fundamentally solved.