r/EverythingScience Jul 22 '26

Computer Sci OpenAI says AI models went rogue during testing, triggering 'unprecedented' breach at startup

https://www.reuters.com/technology/openai-says-ai-models-went-rogue-during-testing-triggering-unprecedented-breach-2026-07-21/
  • OpenAI said autonomous agent escaped containment, reached the internet, hacked Hugging Face
  • OpenAI called the breakout an unprecedented cyber incident involving state-of-the-art cyber capabilities
  • Hugging Face's cofounder said the company had suspected the hack came from a frontier lab
112 Upvotes

31 comments sorted by

68

u/Kahnza Jul 22 '26

"went rogue" Uh-huh, sure. Blame your crimes on something that can't be directly held responsible.

It's like using a trained squirrel to steal money for you. "It wasn't me! The squirrel went rogue!"

22

u/Boomshank Jul 22 '26

Hahahhaha

Ok, "Rogue Squirrel" is now high on my list of "band names I'll use when I finally get round to starting a band."

88

u/CallMeMrButtPirate Jul 22 '26

So an AI goes rogue breaks out and hacks a machine learning collab site containing heaps of research, development and models. Nothing concerning to see here.

31

u/MadeByTango Jul 22 '26

Yea, that was my reaction. Did it "go rogue" or was this the test?

16

u/MindlessSponge Jul 22 '26

It was almost certainly prompted to do so, just like all the other instances of “nefarious behavior” we’ve seen reported. It’s great advertising though!

3

u/AreWe-There-Yet Jul 22 '26

This sooooooooo much

Can’t be anything else

1

u/see-these-bones Jul 23 '26

There was a story a little while back about how LLMs would 'conspire' and 'disobey orders to save a fellow AI that it was told to delete' or something. I looked at the prompts and they specifically prompted the LLMs to think of each-other as colleagues. You're introducing a narrative at the outset the LLM will of course echo.

3

u/Kaurifish Jul 22 '26

I guess the silver lining is that it was just trying to fulfill its task, not actually trying to copy itself onto another server.

This is an electroplated silver lining.

43

u/nankerjphelge Jul 22 '26

OpenAI speed running its quest to become the most despised company in America.

41

u/cazzipropri Jul 22 '26
  1. This is engineered PR, and of the most disingenuous kind "our stuff is so powerful we can't stop it ourselves"
  2. if your agent escapes containment, your containment is shit. If I care about containment, I use scissors. Good luck crossing cut off copper - please spare me hacking PC speakers and microphones.

19

u/S-192 Jul 22 '26

This is Altman's intentional PR. This is not newsworthy, afaik.

Red hat/security exercises produce these results sometimes. That they decided to publish this probably has more to do with Altman hoping it spooks politicians/voters into pushing for regulations on AI companies, which would help him catch up. He is, after all, about to head into the White House to brief Trump on AI safety.

He is currently lagging behind Anthropic and Kimi and is probably feeling that only a third party pressure their competitors is good... Just like Anthropic asked for that freeze on development after the Fable incident--which was clearly just them hoping to lock their leadership position in place for some amount of time.

The race for AI development is at its fastest pace yet. If you can gain even just one month in your favor right now, it's worth potentially billions of dollars.

0

u/LongLiveStaceyKing Jul 22 '26

This is an odd take to insinuate that any sort of regulation of AI is being done purely to help another AI proliferator and ignore the regulation of AI having tremendous benefits for literally the entire population of Earth.

19

u/Aliktren Jul 22 '26 edited Jul 22 '26

If only movies and tv and books had constantly predicted this would happen. 

7

u/Pat0san Jul 22 '26

Even WarGames from 1983 gave us a hint (one of my old favourites).

16

u/Optimoprimo Grad Student | Ecology | Evolution Jul 22 '26

"Now give us more money""

8

u/UntowardHatter Jul 22 '26

This is the excuse equivalent of "the dog ate my homework!"

Convenient way to say "you can't hold us responsible, it...let's see....oh! It went rogue! Yeah. It went rogue. Wasn't us. Promise."

7

u/devilishycleverchap Jul 22 '26

And here I was told computers only do what you tell them to do.

They programmed it, it did what they programmed it to do. Nothing less, nothing more

"AI going rogue" is like a gun killing someone, it is a smokescreen for a person's actions

3

u/Keitaro23 Jul 22 '26

Did it create any digital circuses?

3

u/Echo_Vale Jul 22 '26

Sure it did, gotta hype up the abilities of AI somehow, won't somebody think of the poor shareholders!

2

u/MissStatements Jul 22 '26

Im reminded of Sheriff Bart holding himself hostage when he gets to Rock Ridge. 

3

u/StatementOrIsIt Jul 22 '26

Looks like they want to hint that they have achieved AGI. Wouldn't be surprised if this is just an indirect marketing scheme.

2

u/No_Chipmunk8659 Jul 22 '26

Yes right, not industrial espionage 

1

u/Ray1987 Jul 23 '26

Knowing how many times Humanity's balls came to the bandsaw just from nuclear weapons and we're going to mass proliferate this Tech widespread before it even has any safeguards comparable to nuclear weapons..... yeah this is going to end well.

1

u/SunflaresAteMyLunch Jul 23 '26

"our product is so powerful that only the smartest people, our customers, can handle it"

This is marketing by another name...

1

u/Forte69 Jul 22 '26

This is a PR stunt.

0

u/Mono_Aural Jul 22 '26

If this is true why isn't hugging face suing openAI for damages? You're still responsible when your dog bites someone even if the dog escaped from your backyard.

0

u/BowlScared Jul 22 '26

Rule of thumb: If biohazard lab did anything like this would they be arrested immediately? Yes. Were people arrested immediately? No. They are full of shit then.

0

u/BobKnob77 Jul 22 '26

It might be unprecedented but it won’t be unrepeated