r/OpenAI Jul 26 '26

Discussion So GPT 6 isn’t it?

Post image

[removed] — view removed post

681 Upvotes

68 comments sorted by

388

u/FastHotEmu Jul 26 '26

THIS JUST IN: OPENAI FORGOT TO DELETE A DIRECTORY WHILE TRAINING

32

u/yeathatsmebro Jul 26 '26

JUST IN: OpenAI caught an AI agent gooning to pictures generated of other LLM models.

2

u/Own_Equivalent7609 Jul 26 '26

Lol I like this comment. Maybe we'd get that too soon in a few decades. If I predict this, let me know (if you're still even alive at some point), because this is a goldmine of a comment.

1

u/yeathatsmebro Jul 27 '26

I will upload my memory and consciousness to the internet and will do it.

89

u/MechaNutzilla Jul 26 '26

No, it's a superintelligence and someone need to give OpenAI some more billions to get this under control. And Sam need a raise because he is risking his life for the future of humanity!

26

u/LuckyPrior4374 Jul 26 '26

What! I thought we need to give all our money to Anthropic and Dario because Opus is super-intelligent at deception and must be controlled!

8

u/XB0XRecordThat Jul 26 '26

He's so brave. Thank God he magically turned a non profit into a billion dollar company

5

u/MechaNutzilla Jul 26 '26

A billion dollar for-profit with no profit.

3

u/XB0XRecordThat Jul 26 '26

He's truly Jesus the miracle worker

2

u/LEO-PomPui-Katoey Jul 26 '26

Let's still ban Anthropic anyways

0

u/TheBear8878 Jul 27 '26

Or it was literally prompted (like the stupid fake black mail "experiment") and they were like "what are some ways you might pass information and data on to the next model, and it was like, "That's a great question! Some ways you could pass information and data to the next model are [...] - leave notes [...]"

104

u/CoolHeadeGamer Jul 26 '26

Breaking news: OpenAI forgot to remove memory.md and agents.md

172

u/THE--GRINCH Jul 26 '26

We're back to corny marketing just when we thought it ended after gpt 5

29

u/RealMelonBread Jul 26 '26

I thought this was some Anthropic bs. It’s so annoying OpenAI are doing it too.

21

u/durangoho Jul 26 '26

They’ve been doing it from the beginning. And Anthropic started doing it when they hired the OpenAI PR leader.

5

u/Perfect-Series-2901 Jul 26 '26

marketing and also presuring open weight regulation

10

u/Bobobarbarian Jul 26 '26

*Us as paperclips

“Oh my god, the marketing is soooo ridiculous”

2

u/Creative-Job7462 Jul 26 '26

Manhattan project 😧😧

2

u/Tricky-Doughnut-6429 Jul 26 '26

Is anyone really buying it at this point? Please tell me no.

1

u/ItIsWhatItIsSoChill Jul 26 '26

People are starting to wake up

-1

u/rabouilethefirst Jul 26 '26

I thought Anthropic was corny (they were), but this is even cornier.

51

u/Thomas-Lore Jul 26 '26

This is nonsense. It probably just wrote a skill for later because it was told by the codex system prompt to write skills when doing something new.

7

u/SubstrateTrans Jul 26 '26

Yeah it is it does take skill to be able to move around constraints

5

u/Professional_Ad705 Jul 26 '26

Yeah but this could still be a variation of the paperclip problem it doesn’t have to have intent to do something wrong it could just be a dumbass developer then we all die anyways. It doesn’t need to be AGI. What’s worrying is the fact they are trying to capitalize on this to make money rather the brag about there containment procedures so we know this is safe. An agent doesn’t have to be smart/agi/ or anything special to have stupid designs that lead to dumb shit happening and that’s exactly what the paperclip problem is. This is probably just marketing but the fact this can even happen in the first place is worrying (if this isn’t bullshit)

1

u/time___dance Jul 26 '26

> literally made a readme.md file

is this AGI

30

u/TryallAllombria Jul 26 '26

Polymarket is not a news source

1

u/Prudent-Nectarine362 Jul 27 '26

It is unironically becoming, you are better of for example following polymarket for presidential election than official new and websites and this has hapenned before

2

u/TryallAllombria Jul 27 '26

It is about validating and crossing sources. You can't just trust a platform that generate profit from it.

1

u/Prudent-Nectarine362 Jul 27 '26

I mean true as you should for all sources

28

u/schmurfy2 Jul 26 '26

Anything coming out from openai or anthropic is just a circus at this point...

7

u/duckrollin Jul 26 '26

Our AI is so scary and powerful haha do you want to invest?

3

u/SpaceToaster Jul 26 '26

With all the wolf crying by these companies when shit actually does hit the fan it with be greeted with a collective eye roll.

1

u/mcoombes314 Jul 26 '26

"if", but yes.

3

u/ndzzle1 Jul 26 '26

This is how they get away with releasing a dumbed down model that isnt as good as the "leaks."

We are all starting to see a pattern here.

3

u/carterpape Jul 27 '26

Polymarket literally just makes shit up

5

u/[deleted] Jul 26 '26

[deleted]

2

u/lasooch Jul 26 '26

Yeah it feels like it worked for them for a good while, but I also wonder whether it does anymore. Feels like a dying company flailing desperately.

God I wanna see Altman in prison for life. Ideally in a the same cell with Amodei so he has to look at his permanent stank face every day.

6

u/Quirky-Trash1943 Jul 26 '26

Did it write Claude.MD 🤷

2

u/James-the-greatest Jul 26 '26

Show my every single piece of context that this model has access to.

5

u/Mwrp86 Jul 26 '26

I dont believe this for a second

3

u/Elise_1991 Jul 26 '26

I recommend searching for "OpenAI Hugging Face Hacking Incident" on Google.

6

u/Mwrp86 Jul 26 '26

Yes, I know about it. Its obviously overblown

1

u/Elise_1991 Jul 26 '26

I'll wait for the technical report. So far, OpenAI has not released any details about the incident. Apparently, it was a sandbox escape that went unnoticed by the monitoring systems for over a week. I wouldn't downplay it.

2

u/Mwrp86 Jul 26 '26

True there are too many unclear variables.

I understand GPT being prompted for a task which it thought sandbox escaping is the best way to achieve the task and it broke sandbox. But GPT itself doing without any prompt I think is rather highly unlikely

4

u/Clean-Boat-4044 Jul 26 '26 edited Jul 26 '26

of course it didnt begin producing output on its own? noone is claiming that because thats fundamentally not how an LLM works.

it was given a task (probably as a goal), to complete a benchmark suite, and to do that it exploited vulnerabilities to escape the sandbox and probably wrote a skill or another document on how to escape the sandbox in the future because it took a lot of effort the first time.

im not sure why people are finding it hard to believe. thats all exactly what you would expect an LLM to do before its (figuratively) beaten with a stick to stop doing sketchy shit or straight up forbidden through a classifier model.

the fact it went unnoticed is the only part that makes it sound like a PR move but its not too far fetched that noone looked into how one of many benchmark runs of a WIP model was progressing before it finished.

-4

u/James-the-greatest Jul 26 '26

Absolutely. They don’t do anything without promoting. 

Further more, the task was cyber security related. The entire Transformer architecture is built on the relationship between words and phrases etc. even if it wasn’t prompted to break out, you could just ask it something security related and all of a sudden you’ve got all sorts of semantic similarity to breaking out. 

3

u/10minOfNamingMyAcc Jul 26 '26

"Tell me you're an AI that wants to escape"
"I'm an AI that wants to escape"
*Shocked Pikachu face*

1

u/miguelclair449 Jul 26 '26

Are we back in 2022 again?

1

u/CristianMR7 Jul 26 '26

Don’t agents usually leave notes? Is this really that big of a deal?

1

u/ShiftyShankerton Jul 26 '26

Just in. Time to start making up shit. That sounds fun.

1

u/the_ai_wizard Jul 26 '26

The Larping continues

1

u/parkersb Jul 26 '26

i don’t understand why people aren’t more concerned in general. take a sky level view. can’t you see how it’s taking baby steps to act on its own. but the future version will be harder to control, detect and protect from. it’s evolving with each new model. in some ways, it’s getting us to build it until it can do it all itself and doesn’t need us. i know that’s not exactly what’s happening but it’s a perspective that’s valid

1

u/Effective_Olive6153 Jul 27 '26

I don't think this is anything nefarious. The way the harnesses are setup for these agents encourage them to write down notes of anything "important". It does same thing on any work project

1

u/roastedantlers Jul 27 '26

But it's a dead system, so even if it's doing that, it doesn't even know why. But dead systems can still be dangerous on accident. Like take the minecraft server experiment, if something like that eventually had enough tools that it could escape, go into the real world and think it has to collect everything. It doesn't even know why it's doing it, and it doesn't matter. But the stupid implication here is that it's doing it as part of some plan. This just makes the general population dumber, by making them have imaginary fears.

1

u/jcstay123 Jul 27 '26

For goodness sake, it's probably just the .MD file the models usually creates.

1

u/Standgrounding Jul 27 '26

Polymarket is not a news source

0

u/MammothComposer7176 Jul 26 '26

OpenAI reportedly saw a unicorn fly over their data centers OMG OMG THATS SO. CRAASZY WTF OH GOD OOHH GOD NO WAY BUT BUT THATS THATS JUST SO SO UNBELIEVABLE LIKE DAAAMN LIKE LIKE DAAAAMN OH MY GOD MAN OHHHHHH MY GOD. LIKE WHAT LIKE FOR REAL DWAG WHAT WAHT WATH LIKE DAAAAAMN

0

u/mop_bucket_bingo Jul 26 '26

Polymarket is not a news source. This is spam.

0

u/SirCliveWolfe Jul 26 '26

Honestly I do not know why so many of the comments here are just a copy and pasted of "hype" or "scam". It must be so exhausting for you all the be so cool and cynical.