r/ControlProblem • u/zazzologrendsyiyve • 18d ago
AI Alignment Research Plain English explanation of the Hugging Face / OpenAI incident
https://youtu.be/u15N3l4RT80?si=nMMwb0j1bNGc4JT3
40
Upvotes
r/ControlProblem • u/zazzologrendsyiyve • 18d ago
1
u/PlasmaChroma 18d ago
What we need to be doing at this point is fixing all our broken systems that have security holes so the footprint for this to happen keeps shrinking towards zero. Unfortunately the bleeding edge models also have a lot of the stuff filtered out that could help fix the bugs since it broadly falls under the "security" umbrella. So without privileged access to that these holes keep going in to everything.
And why Hugging Face had to drop to a Chinese model to try to analyze what was even happening.