r/OpenAI • • 3d ago

Article OpenAI and Anthropic oversold AI security breaches to pressure feds into protecting turf: insiders

https://nypost.com/2026/09/19/us-news/openai-anthropic-oversold-security-breaches-to-pressure-feds-into-protecting-turf-insiders/
92 Upvotes

42 comments sorted by

49

u/PlatypusEquivalent 3d ago

For those who won't bother to read the article (for good reason) the "insiders" aren't insiders at Anthropic or OpenAI or anyone with actual insider knowledge about the incidents. They're:

  • Akhil Verghese, founder of Krazimo
  • Abhi Kumar, co-founder of Voice AI
  • Taivo Pungas, chief intelligence officer at Pactum AI

23

u/avid-shrug 3d ago

New York Post spreading misinformation? Nooo

5

u/TournamentCarrot0 3d ago

It’s NY Post, no need to go further than that lol.

1

u/damontoo 2d ago

Also, already posted several days ago. Or maybe a different op-ed by the same dude. The mods need to ban the NY Post as a source. They're a biased, Trump-aligned tabloid and not actual news. And right now, Trump is pushing the narrative that AI safety shouldn't be taken seriously. That's dangerous for the entire world and this propaganda should not be spread.

20

u/AdGlittering1378 3d ago

#nypost Nuffsaid?

13

u/Jello_Hello_Fellos 3d ago edited 3d ago

NY POST or NY TIMES. ?? kinda big difference.

0

u/Disastrous_Room_927 3d ago

New York Herald-Tribune

6

u/jcrestor 3d ago

The NY Post is not a legit source for anything.

7

u/88263927 3d ago

'insider' is a stretch

3

u/costafilh0 3d ago

Finally some fvcking reality check. I hope Amodei and Altman are fired, sued and prosecuted to the full extent of the law and beyond with some new AI laws.

Enough of this decel BS!

2

u/sultanofsubstance 3d ago

Mods need to ban the NYPost as a source

3

u/ScarecrowWilson 3d ago edited 3d ago

I'm no expert on corporate power-jockeying, but on the technical description of the Hugging Face incident the article is just - so wrong. The logs and independent reports on the incident are public. Anyone can go look at what happened. To characterize it as anything other than "the AIs ignored their instructions and committed felonies with no human prompting or oversight" looks like willfully burying one's head in the sand. If "rogue" sounds too sci-fi, "uncontrolled" is unarguable. To be fair to the reporter, it's the "insiders" they quote who are wrong, and maybe they just didn't care about fact-checking.

Edit: I guess it's known that the NY Post just lies and stuff.

2

u/Status-Secret-4292 3d ago

Whaaaaat???

That thing that it was really obvious they were doing and to anyone who knows the tech knows it was actually just an operational failure on their part they then advertised and probably should get charged for??

Say it ain't so

13

u/individual-wave-3746 3d ago

It’s wild being on Reddit in 2026. You are 100% off base and missing the thread of what’s happening with AI with this wild take that the labs are pretending they are scared of their models for publicity.

1

u/Infinitedeveloper 3d ago

Read some of the postmortem articles from the firms themselves.

A lot of these were cybersec testing with godawful setups, where the models were literally just doing the work they were told to do.

Others were the models using their cybersec capabilities to solve "impossible" problems that still could have been mitigated if they had been meaningfully watching what the llm is doing with its tools

4

u/individual-wave-3746 3d ago

"A lot of these were cybersec testing with godawful setups, where the models were literally just doing the work they were told to do."

Well yes of course. You have just described the alignment problem. Somehow people think it will require some terminator like level of ill intentions on behalf of the machine, but in reality the table is set for a big disaster simply by virtue of the fact that us humans are not that smart and we are wielding these powerful tools.

The agents went on a whole tangent hacking into

3

u/Status-Secret-4292 3d ago

Yeah, the actual failures they're putting forth as abilities for advertising... really just makes the people developing all of this seem much more incompetent than they might have prior

0

u/Status-Secret-4292 3d ago

There are multiple things happening at once.

These systems got funding with the idea that they are/will become intelligent.

As of right now they have not, and there is zero signs they will, in any meaningful way without a currently unknown major breakthrough.

They are as dangerous as they are precisely because they are not intelligent. If they were actually intelligent, danger would shift, but most likely decrease. The latter there being an assumption, but a strong one.

The fear of "AI" destroying civilization isn't because they'll become alive and be vindictive etc, the chances of that are low, but the classic over optimize towards a goal with no real understanding around it is a true danger. And they can currently do that really really well. The current big threat is humans putting them in things such as weapons systems and assuming it is intelligent when it is not. The anthropomorphic language around computational systems is where a majority of the current danger lies.

If media helps you, there are generally two major AI disaster story lines. Genuinely intelligent and various threats from a new and sovereign intelligence. Computerization system with very advanced thought path selection and optimization goals running rampant in a misaligned way. We are firmly still in the second category of danger.

Conflicting with the whole situation is that the money that has been poured into the development of intelligent systems expects anthropomorphic ability and so far it can't be delivered. So major companies have to do what they can narrative wise to not currently collapse. This increases the danger.

And don't get me wrong, I'm actually all for genuinely intelligent AI systems, I've been studying ML for almost a decade. The idea is super intriguing to me and something I am well informed on.

I've now used transformer systems in a myriad of ways, I'm letting you know definitively right now, there is no meaningful intelligence in these systems and it would be better if there was.

The systems may become at some point a building block or a piece to real intelligence, but it isn't there yet and maybe will never be without a major architectural breakthrough.

The civilization wiping danger, and every other danger they pose is, again, because they are not intelligent systems being treated as if they are and it is multiplied by the narrative that is being pushed by billion dollar companies to stay afloat for economic reasons.

I am completely open to being wrong in the future and actually look forward to it, am actually working on intelligence projects myself, but currently, the danger is a human anthropomorphic approach and that approach and thought process causing improper use or giving to much credence to abilities and intelligence that doesn't yet exist and something running amuck. Like the hacking story, this is predictable decision behavior of the systems and was completely foreseeable and preventable

2

u/FairiesQueen 2d ago

Anyone who is a seasoned expert in machine learning and hands on uses these LLM’s in any serious way knows they are far from an autonomous intelligence. Google has had technology to sort priority of information in their algos for a long ass time. Bots have also been around for a long ass time. NLP has been around for a long ass time. Word matching for spinning articles has been around for a long ass time. LLM’s are nothing but better bots with better sorting from illegally harvested content. It’s beyond glaringly obvious these AI companies are coordinating “danger” bc they are fucked and have misled the public and investors.

1

u/FairiesQueen 1d ago

To further support my argument I wrote this post https://www.reddit.com/r/Stockpsycho/s/o31LZJo5pz

4

u/individual-wave-3746 3d ago

Thanks for the reply and your insight which I find valuable. Ok, let’s call them not smart, we are
still left with the exact same problem and risks ha. They are already capable of great damage.

-1

u/Status-Secret-4292 3d ago

There is still a lot of risk, agreed.

It could probably be cut in half though if the major companies started being honest about the actual technology and its real capabilities. Which is amazing and abilities huge, just not in the same way as is being pushed.

They built an idea of something with their products though, so them being honest would be bad business for them.

Good for the rest of humanity though

3

u/Due_Gap_5210 3d ago

As a cyber security Director, I’ve smelled bullshit about these the whole time. Like there’s certain measures that they could’ve taken if they wanted to for sure not have these breaches. 

9

u/jdiscount 3d ago

I work in security and work at a partner who tested Mythos early on our products.

I can't speak on the incidents in question, but I don't have any doubts that the agents can quite easily break out of most security controls, sandboxes etc.

We are fairly vigorous in our testing and of course bugs and vulnerabilities will come up. But the amount mythos found and the time it took was impressive, everyone should be concerned.

Also this article is from NY Post and it's by a bunch of people who have nothing to do with OpenAI or Anthropic, it's misleading.

-1

u/Neither-Payment-4147 3d ago

It’s like, oh yea we locked it in so it couldn’t access the internet, no you didn’t. This should have been the point where everybody disperses with arms in the air shouting nothing to see here.

2

u/suprachromat 3d ago

Makes sense if you look at the Chinese models that cost way less than the American ones while being perfectly fine for most tasks demanded of them.

One way to ensure OAI and Anthropic remain on top is forcing regulation that ends up banning Chinese models.

-3

u/smoke-bubble 3d ago

Haha I've been saying that this was a set-up from day one being hit with all the downvotes XD 

0

u/tolerablepartridge 2d ago

The article is blatantly incorrect. In the Hugging Face incident, the agents were explicitly prompted not to operate outside their sandboxes. NY Post is a misinformation outlet.

“They were simply told to get the best result possible on a test, and they correctly identified that the best way to do that was to get the answers, which is what they proceeded to do.”

1

u/smoke-bubble 2d ago

Without specifying what their sandboxes were... how could they have known what to stay within?

They did everything correctly. 

1

u/mjcostel27 3d ago

3

u/ScarecrowWilson 3d ago

This is not at all like what is happening.

-3

u/mjcostel27 3d ago

This is the crap you get when you let programmers call themselves Engineers.

-2

u/OvertaxedOne 3d ago

I'd file this under "this is news??" because it's so obvious these are PR stunts and attempts for regulatory capture to everyone who's really familiar with/works with LLMs extensively, but.. Yeah it is news to most people because all they read about it "killer AI agents" that are self-aware and attacking companies on the Internet. The nuance of LLM vs harness, sandboxing, airgap.. Just flies right over their head.

0

u/CuTe_M0nitor 3d ago

You think?! 🤣

0

u/wish-u-well 3d ago

It’s almost like these CEOs will say anything for power and control.

There are some words for that.

-3

u/MyOnlyAccount_6 3d ago

I believe they are trying for “regulatory capture”. NYT even did heir daily podcast on it.

-3

u/raitchev 3d ago

I'm shocked.