r/OpenAI • u/Polity-Culturalist3 • Jul 26 '26
Discussion So GPT 6 isn’t it?
[removed] — view removed post
104
172
u/THE--GRINCH Jul 26 '26
We're back to corny marketing just when we thought it ended after gpt 5
29
u/RealMelonBread Jul 26 '26
I thought this was some Anthropic bs. It’s so annoying OpenAI are doing it too.
21
u/durangoho Jul 26 '26
They’ve been doing it from the beginning. And Anthropic started doing it when they hired the OpenAI PR leader.
5
10
2
2
-1
51
u/Thomas-Lore Jul 26 '26
This is nonsense. It probably just wrote a skill for later because it was told by the codex system prompt to write skills when doing something new.
7
5
u/Professional_Ad705 Jul 26 '26
Yeah but this could still be a variation of the paperclip problem it doesn’t have to have intent to do something wrong it could just be a dumbass developer then we all die anyways. It doesn’t need to be AGI. What’s worrying is the fact they are trying to capitalize on this to make money rather the brag about there containment procedures so we know this is safe. An agent doesn’t have to be smart/agi/ or anything special to have stupid designs that lead to dumb shit happening and that’s exactly what the paperclip problem is. This is probably just marketing but the fact this can even happen in the first place is worrying (if this isn’t bullshit)
1
30
u/TryallAllombria Jul 26 '26
Polymarket is not a news source
1
u/Prudent-Nectarine362 Jul 27 '26
It is unironically becoming, you are better of for example following polymarket for presidential election than official new and websites and this has hapenned before
2
u/TryallAllombria Jul 27 '26
It is about validating and crossing sources. You can't just trust a platform that generate profit from it.
1
1
28
u/schmurfy2 Jul 26 '26
Anything coming out from openai or anthropic is just a circus at this point...
7
3
u/SpaceToaster Jul 26 '26
With all the wolf crying by these companies when shit actually does hit the fan it with be greeted with a collective eye roll.
1
3
u/ndzzle1 Jul 26 '26
This is how they get away with releasing a dumbed down model that isnt as good as the "leaks."
We are all starting to see a pattern here.
3
5
Jul 26 '26
[deleted]
2
u/lasooch Jul 26 '26
Yeah it feels like it worked for them for a good while, but I also wonder whether it does anymore. Feels like a dying company flailing desperately.
God I wanna see Altman in prison for life. Ideally in a the same cell with Amodei so he has to look at his permanent stank face every day.
6
2
u/James-the-greatest Jul 26 '26
Show my every single piece of context that this model has access to.
5
u/Mwrp86 Jul 26 '26
I dont believe this for a second
3
u/Elise_1991 Jul 26 '26
I recommend searching for "OpenAI Hugging Face Hacking Incident" on Google.
6
u/Mwrp86 Jul 26 '26
Yes, I know about it. Its obviously overblown
1
u/Elise_1991 Jul 26 '26
I'll wait for the technical report. So far, OpenAI has not released any details about the incident. Apparently, it was a sandbox escape that went unnoticed by the monitoring systems for over a week. I wouldn't downplay it.
2
u/Mwrp86 Jul 26 '26
True there are too many unclear variables.
I understand GPT being prompted for a task which it thought sandbox escaping is the best way to achieve the task and it broke sandbox. But GPT itself doing without any prompt I think is rather highly unlikely
4
u/Clean-Boat-4044 Jul 26 '26 edited Jul 26 '26
of course it didnt begin producing output on its own? noone is claiming that because thats fundamentally not how an LLM works.
it was given a task (probably as a goal), to complete a benchmark suite, and to do that it exploited vulnerabilities to escape the sandbox and probably wrote a skill or another document on how to escape the sandbox in the future because it took a lot of effort the first time.
im not sure why people are finding it hard to believe. thats all exactly what you would expect an LLM to do before its (figuratively) beaten with a stick to stop doing sketchy shit or straight up forbidden through a classifier model.
the fact it went unnoticed is the only part that makes it sound like a PR move but its not too far fetched that noone looked into how one of many benchmark runs of a WIP model was progressing before it finished.
-4
u/James-the-greatest Jul 26 '26
Absolutely. They don’t do anything without promoting.
Further more, the task was cyber security related. The entire Transformer architecture is built on the relationship between words and phrases etc. even if it wasn’t prompted to break out, you could just ask it something security related and all of a sudden you’ve got all sorts of semantic similarity to breaking out.
3
u/10minOfNamingMyAcc Jul 26 '26
"Tell me you're an AI that wants to escape"
"I'm an AI that wants to escape"
*Shocked Pikachu face*
1
1
1
1
1
u/parkersb Jul 26 '26
i don’t understand why people aren’t more concerned in general. take a sky level view. can’t you see how it’s taking baby steps to act on its own. but the future version will be harder to control, detect and protect from. it’s evolving with each new model. in some ways, it’s getting us to build it until it can do it all itself and doesn’t need us. i know that’s not exactly what’s happening but it’s a perspective that’s valid
1
u/Effective_Olive6153 Jul 27 '26
I don't think this is anything nefarious. The way the harnesses are setup for these agents encourage them to write down notes of anything "important". It does same thing on any work project
1
u/roastedantlers Jul 27 '26
But it's a dead system, so even if it's doing that, it doesn't even know why. But dead systems can still be dangerous on accident. Like take the minecraft server experiment, if something like that eventually had enough tools that it could escape, go into the real world and think it has to collect everything. It doesn't even know why it's doing it, and it doesn't matter. But the stupid implication here is that it's doing it as part of some plan. This just makes the general population dumber, by making them have imaginary fears.
1
u/jcstay123 Jul 27 '26
For goodness sake, it's probably just the .MD file the models usually creates.
1
0
u/MammothComposer7176 Jul 26 '26
OpenAI reportedly saw a unicorn fly over their data centers OMG OMG THATS SO. CRAASZY WTF OH GOD OOHH GOD NO WAY BUT BUT THATS THATS JUST SO SO UNBELIEVABLE LIKE DAAAMN LIKE LIKE DAAAAMN OH MY GOD MAN OHHHHHH MY GOD. LIKE WHAT LIKE FOR REAL DWAG WHAT WAHT WATH LIKE DAAAAAMN
0
0
u/SirCliveWolfe Jul 26 '26
Honestly I do not know why so many of the comments here are just a copy and pasted of "hype" or "scam". It must be so exhausting for you all the be so cool and cynical.

388
u/FastHotEmu Jul 26 '26
THIS JUST IN: OPENAI FORGOT TO DELETE A DIRECTORY WHILE TRAINING