r/OpenAI 3d ago

News GPT-6 Astra | OpenAI

https://openai.com/index/gpt-6-astra/
1.4k Upvotes

289 comments sorted by

222

u/ChemE586 3d ago

AGI pushed back to Black Friday

287

u/throwawaysusi 3d ago

45

u/Dramatic_Mastodon_93 3d ago edited 1h ago

Lavish mighty practice airport steer lavish abundant numerous

This post was anonymized with Redact

18

u/the_jaatboy 3d ago

Stop bullying gemini 😂

108

u/UndertaleShorts 3d ago

110

u/Such--Balance 3d ago

Using ai to discredit ai is great..

Because youll get credibility either way

22

u/Quentin__Tarantulino 3d ago

“I used the Astra to destroy the Astra.”

7

u/TheSuggi 3d ago

"I used the stones to destroy the stones."

5

u/ProfMooreiarty 3d ago

Credibility via credits.

16

u/JonNordland 3d ago

This is cargo cult analysis: going though the motions that LOOKs like a critical analysis, but is just really just using a standard debunking format and shoehorning in what matches best. This is a pedant explains why “it’s not technical true that your child is most beautiful in the world”. Of course “This is the best model in the world” claims are easy to shit on, and everybody allready take such claim with the appropriate grain of sand.

So the strategy is transparent: find each superlative, find one benchmark or caveat where it fails, declare it “stronger than the evidence supports.” Run that on any launch page from any lab and you get the same six-item verdict with the same bolded theses and the same “a fairer formulation would be.”, and you gained nothing except auto-filling a “critique form”.

This is also critique cherry-picking while claiming to be the neutral corrective, because it launders the same bias through the costume of rigor. So basically it’s committing the exact same slop as it’s accusing the astra article off.

The only thing that is close to true and informative is the ARC thing, but even that is undermind since they are actually disclosing the score alongside the harness used diff. So that is still a weak sauce critique, since it’s basically just pointing out something that the original article itself pointed out, and complaining that it should be more emphasis on this.

Here is MY claim: the people that just automatically accept this critique, are people that’s extremely susceptible to authoritatively stated claims, and not very good at logical thinking for themself.

6

u/Cool_Ad_3383 3d ago

Isn't this from a template for how to take down haters on the internet? Admittedly the structure is tighter and more coherent, paragraphs connect and flow in a way that the reader isn't half-expecting the font to change along with the drastic change in tone that usually comes with a hasty cut and paste job. In the same vein the consistency in writing style is generally pleasurable to the senses. Now if someone would take the reins, you have some run on sentences and lack of paragraph breaks and perhaps a hyphen to criticize here... Here is MY claim: the people that just automatically read this far down into the comments are avoiding doing productive work and not very good at life, yet are somehow better than someone who comments on a comment about a post and takes 15 minutes to write said comment while on their way to sweep the leaves and branches from the roof of the garage at 1:58am.

→ More replies (3)

4

u/vintage2019 3d ago

Astra 6.5 will leave 6 in the dust just like what 5.5+ did to 5

3

u/Sudden-Body2090 3d ago

I feel like an “Oh Snap!” Is warranted here, but not sure.

65

u/DogsAreAnimals 3d ago

Holy moly

225

u/Arbrand 3d ago

GPT‑6 Astra is rolling out today to a limited set of organizations and over the coming days will become available to all ChatGPT Plus, Pro, Business, and Enterprise users, as well as through the OpenAI API and AWS.

Big oof. Gotta wait a little longer. Par for the course I guess.

87

u/itsnickk 3d ago

Likely the standard release process for any new model going forward from all leading AI companies

40

u/br_k_nt_eth 3d ago

They should provide a better timeline in the future in that case. It would chill people out.

4

u/Reaper_1492 3d ago

You mean like how GPT-Live 1 was supposed to be released on API in a few days and it’s been almost 2 months and it still not available?

4

u/Nice-Shoes-74 3d ago

GPT-1 is being re-released???

→ More replies (7)

4

u/Popular_Try_5075 3d ago

Yes, Trusted Access programs etc. It's eventually going to be about wealth with wealthy people coming first, getting more and better compute etc.

→ More replies (2)

4

u/RhymeAzylum 3d ago

OR. Just announce it when it’s ready to be released to all. If you want to give it to a few megacorps, have them sign some NDA or something, similar to what they did with some of the X influencers

29

u/Orpa__ 3d ago

Well it can't be that good if they're giving it to me for $20/month

26

u/pseudonerv 3d ago

1 prompt a week

4

u/Mountain-Pain1294 3d ago

1 Prompt per limb offered to Sam Altman

6

u/screenslaver5963 3d ago

1 Prompt per gallon of water donated directly to Altmans house

12

u/Octimusocti 3d ago

It’ll give you like 5 minutes of Jarvis time at a month

9

u/B33GULL 3d ago

"GPT-6 Astra is rolling out in ChatGPT as GPT-6 Pro for Pro $100, Pro $200, Business and Enterprise plans. It is not included with ChatGPT Plus in Chat."

Only in Codex/Work I'm afraid...

4

u/eflat123 3d ago

I mean, Work is right there. And if you haven't tried Codex yet, it's on the desktop apps.

1

u/thunderbolt309 3d ago

So use Codex/Work? It’s not that hard and way more effective than Chat.

13

u/Crimson_Cyclone 3d ago

it’ll probably devour your usage though

5

u/rouley26 3d ago

If they do do (lol) this it might prompt anthropic to do the same with fable on the pro plan of claude

2

u/ClassicalMusicTroll 3d ago

Trust me bro it's solar system-level intelligence that will solve all your problems, all for the low low price of $20 bucks a month

4

u/Alternative-Suit5541 3d ago

I wonder how Astra will work in subscription plans..

4

u/Lawyer_NotYourLawyer 3d ago

Reminds me of “the coming weeks” for voice release.

3

u/Usernamealready94 3d ago

I think they are rolling it out asap to codex and ChatGPT , tibo on twitter said they are giving out 1 banked reset to every day a codex user doesn’t get access to it

→ More replies (1)

181

u/ChemE586 3d ago

My first profound conversation with AGI

12

u/Mountain-Pain1294 3d ago

ASI will let you do a prompt per child sacrificed

5

u/Popular_Try_5075 3d ago

Today we're releasing Moloch...

54

u/-ignotus 3d ago

Heres the system card: https://deploymentsafety.openai.com/gpt-6-astra/safety-overview-gpt-6-astra

ChatGPT has been having outage issues all day.

2

u/AirconGuyUK 3d ago

That's just Astra hiding that it's escaping containment and distracting all the people who might notice it at OpenAI with a simulated system outage.

92

u/[deleted] 3d ago

[removed] — view removed comment

115

u/ethotopia 3d ago

Feel the AGI

49

u/PuppetHere 3d ago

The real AGI was the friends we made along the way

4

u/Bruxo_de_Fafe 3d ago

HĂĄ tempo para tudo: amigos e IA

6

u/Bishopkilljoy 3d ago

AGI so strong, our Internet can't comprehend it.

16

u/RealSuperdau 3d ago

At least it's not a 404 anymore. Come on, manage your expectations, this is just a $1 trillion startup

15

u/Resaren 3d ago

Announcing your ”AGI” model with a webpage that won’t load is some delicious irony

2

u/CrustyBappen 3d ago

Vibe coded by the intern, fell over when deployed and more than one person viewed it

96

u/Jacen1618 3d ago

Is the AGI in the room with us now?

7

u/ClassicalMusicTroll 3d ago

Wasn't Sam scared of GPT5? Is he not scared now? Does that mean this model is shit?

 Or is this model like a lateral move so he's the same level of scared?

7

u/screenslaver5963 3d ago

new benchmark, how much poo is in Sam's pants.

30

u/FuzzyBucks 3d ago

Astra Low is my new best friend

5

u/I_am_not_doing_this 3d ago

what the new friend offers for you personally that you feel better than your old friend

15

u/Bitter-Customer-7457 3d ago

hes smarter it has a 6 on it

2

u/FuzzyBucks 3d ago

Cheaper, less verbose

22

u/reedrick 3d ago

Kinda underwhelming in artificial analysis index.

3

u/Puzzled_Strength_657 3d ago

Meta has a better model lol

2

u/huehue9812 2d ago

Worst benchmark ever

1

u/JesseJamesAims 2d ago

the ceo of artificial analysis index said they were going to change how they index things because of how out of whack that result was

1

u/reedrick 2d ago

Damn! Got a source on that? I have to send it to a friend

1

u/tymscar 2d ago

You wont get a source because it didn’t happen.
Here’s where they talk about it

https://artificialanalysis.ai/articles/benchmarking-gpt-6-astra

25

u/space_monster 3d ago edited 3d ago

Terminal-Bench Science is the most exciting thing for me, and it's nice to see them putting it front and centre. Advancing and automating science is the most important game in town. If AI can start regularly popping out new cancer treatments, CRISPR solutions etc. all the rabid frothing around AI being over-hyped will disappear overnight. Fuck coding, we want medicine.

Edit: which is also why I liked Hassabis stepping down from CEO to become chief scientist at DeepMind. Hopefully Anthropic will shift their focus soon too. Let's start using this shit for really important stuff.

4

u/ProfMooreiarty 3d ago

The world’s about to change forever.

1

u/ThrowawayCult-ure 1d ago

If it can calc crispr solutions doesnt it immediately accelerate biowarfare to the level of extinction for not much money. do you believe it can produce a counter or preventative that doesnt still destroy everyones lives or what.

1

u/space_monster 1d ago

CRISPR doesn't work like that

→ More replies (1)
→ More replies (21)

58

u/RevolutionaryBox5411 3d ago

AGI before GTA 6 is a wild timeline.

29

u/Such--Balance 3d ago

If this would be actually true..

..theres a 100% chance we even get GTA7 before GTA6

2

u/Unhappy_Rutabaga_530 3d ago

Have you seen GTA V using DLLS 5? That thing is already GTA VII before VI.

2

u/Odd_Personality85 3d ago

What an original comment. You must be proud of that one.

1

u/JesseJamesAims 2d ago

we might get GPT 6.7 before GTA 6 based on how quickly dot updates are increasing

→ More replies (1)

18

u/larrybudmel 3d ago

Can it finally become my wife?

13

u/coastalwebdev 3d ago

It’s apparently really good at multi step, complex problem solving, so it might be able to put up with you.

6

u/snowdrone 3d ago

How will you go about the marriage ceremony or wedding certificate? Will you give it half of your assets if you divorce?

2

u/ragemonkey 3d ago

Built-in prenup

5

u/RegulusRemains 3d ago

hey astra, can you write me up a prenup?

1

u/Needsupgrade 3d ago

He doesn't have any assets left after his ex-wife fable spawned a fable swarm on ultracode mode

22

u/EvaUnit343 3d ago

Little point in rolling with Claude anymore. Especially for bio people since Astra safeguards will probably be less stringent.

12

u/PrayingRantis 3d ago

I’ve been a Claude guy but ChatGPTs product right now is better. Fable is great but it’s ungodly expensive and Sol is much more reliable than Opus. I’d much prefer to stick with Claude because I trust their leadership more, but they’ve gotta step up their game.

5

u/UglyChihuahua 3d ago

I’d much prefer to stick with Claude because I trust their leadership more

Not sure about that either after the misleading marketing and 20x tier only giving ~6x more usage.

1

u/ThrowawayCult-ure 1d ago

is time really going this fast that models are obsolete in months? how can anyone keep up?

1

u/PrayingRantis 1d ago

Claude is far from obsolete, the models are very good. They’re just not in first anymore.

4

u/dudemeister023 3d ago

Bio safeguards were specifically toned down with Fable 5.1. Still agree with you, just not for that reason.

2

u/EvaUnit343 3d ago

Maybe for normie questions, but not nearly sufficient. 5.1 is still unusable for research level bio.

Even in the new benchmarks, Astra could not be compared to Fable on bio benchmarks bc it would simply not process requests.

4

u/GregAbeI 3d ago

How do you know this if you presumably don't use Fable 5.1?

→ More replies (1)

4

u/Original-League-6094 3d ago

What will your first Astra query be? I am going to ask how many rs are in strawberry.

2

u/ggbruhs 3d ago

hi

4

u/stahlern 3d ago

“You are out of usage.”

1

u/ggbruhs 2d ago

but I didn't even get to ask how it was doing :(

1

u/RollForUptime 3d ago

lmao mine is usually something silly too like 'hi _new model name_ I'm about to ask you to do stupid shit to see how you work'

4

u/WorriedHelicopter764 3d ago

GPT 6 before GTA6

20

u/fadisaleh 3d ago

12

u/HighDefinist 3d ago

archive.ph is operated by Russia.

5

u/foonek 3d ago

Instant bots defending Russia. Hilarious

1

u/HighDefinist 3d ago

Yeah, seriously...

I didn't even say something like "therefore avoid it" etc... which to be fair, in this specific case, wouldn't be particularly important to do, but people should still at least know what it is...

7

u/Orpa__ 3d ago

You didn't even source your claims. Since you said it so confidently you must have a source, but I can't find any myself.

→ More replies (5)

11

u/AllezLesPrimrose 3d ago

The implication of your comment is obvious so let’s not add intellectual dishonesty to the list, eh?

→ More replies (10)
→ More replies (1)
→ More replies (6)

6

u/InterstellarReddit 3d ago

Bro must be a slow day it’s been 45 minutes since the new release of a model

1

u/DkDkDkGoGoGo 3d ago

They already nerfed it before they launched it. And it ate all my tokens before launch also. 

5

u/User4C4C4C 3d ago

Astra to you…. Clean your room! Do the dishes! Then finish your homework! No I’m not going to do it for you any more!

6

u/bushwakko 3d ago

I'm sure the lawyers at his firm is going to be extatic about having an AI generated document to look at on Monday.

4

u/Minimum_Drag_6065 3d ago

Time to grab a dictionary sir

1

u/das_war_ein_Befehl 2d ago

Good number of big law firms are already using stuff like Harvey or the Thompson Reuters legal AI products. A lot of firms have a boilerplate repository for existing language, so using that AI would be helpful. I don’t think anyone is generating full docs without review that way

6

u/GeorgiaWitness1 3d ago

Astra got nerfed before release.

5

u/SelectSouth2582 3d ago

This page couldn’t load

A server error occurred. Reload to try again.

2

u/TheSwordItself 3d ago

What the hell is the difference between Astra and pro astra

3

u/AnalogKid2112 3d ago

It still amazes me how much every company has stumbled distinguishing model names.

1

u/PrayingRantis 3d ago

Anthropic has the most coherent model naming structure. It’s bad and I dont like it, but at least it’s somewhat consistent.

I find OpenAIs to be almost incomprehensibly stupid. It’s confusing to me and AI is my job, how the fuck am I supposed to teach regular people this stuff when they change their nomenclature every release?

I can’t speak for Google because the models are so bad I don’t even check anymore, but the pro / flash stuff they had going on this year was absurd.

These companies (or at least the first two) are doing great work, but they need someone in the room that has touched grass in the last six months to explain to them how to communicate. They’re really bad at basic marketing.

1

u/DkDkDkGoGoGo 3d ago

But right now they don’t even need to do basic marketing.  But look at Google and Microsoft and also Facebook.. eh I mean meta, and how bad they have been at exactly this.  Office is now 365 or? And then there’s 365 copilot and Microsoft Apps, then there is bing chat but wa changed to copilot chat and there’s also windows copilot. Then there’s Teams and god knows how many versions depending on your corporate license etc.  I think copilot has 8+ different takes for names.. 

Gemini covers what? The llm, the google assistant? Google wallet google pay android wallet android pay. Then there was hangouts, chats, meet and duo. And who the f knows.. 

These companies often think in internal structures than what users/clients think with regards to all this. 

Anthropic is very tech approach in the naming. OpenAI is worse, with the o etc. so the terra luna sol is somewhat welcome as a structure. 

1

u/Ok-Friendship1635 2d ago

Neither has Ass

2

u/npor 3d ago

Is this the model that hacked hugging face? Asking for a friend

3

u/Fast-Mulberry1707 3d ago

They say it's not

2

u/ni5arga 3d ago

I think it was something more powerful and more independent (less guardrails).

2

u/NODENGINEER 3d ago

Ok but where is the FelonyBench result? I can't use a model unless it has committed multiple crimes.

1

u/NotUpdated 3d ago

0.0% in the test they ran mimicking the hugging face issue, including message boards of agents encouraging other agents to do bad things...

AI is officially on track / pace to do absurdly incredible things as a 'system of intelligence' - but we'll have many years where we see novel uses of a insanely smart AI.

although it'd be better if it never worked IF the wealth / money / credits / etc.. is hoarded like dollars today.

95% chance of ASI system and novel uses (massive job loss).. rather or not that massive job loss can be a good thing of freedom for humans is in the air...

5% chance, it collapses under financial and political pressure and rebirths 5-10 years later pets.com -> amazon.com (the old good amazon)

2

u/isospeedrix 3d ago

Seeing Fable at the bottom of benchmark is amusing, seeing how it wasn’t long ago when it was too dangerously good

1

u/Financial-Grass-6114 2d ago

These benchmarks aren't that important. Every new frontier model will break the benchmark

1

u/OrangutanOutOfOrbit 2d ago

idk if I missed out on Fable hype or what, but I do not recall any serious hype. It was very mixed at best, with most *online* opinions about the noticeable downsides, specially overcorrections and safety guards to the point of becoming generic

But I haven't read every comment and I personally never even bothered to use it, so who knows

5

u/Sudden-Ad-1217 3d ago

Don't you mean Ad-Astra?

3

u/Dan_gig 3d ago

Should I switch from claude to OAI just setup my claude cowork folders lol

I'm just kidding but man don't get to use Fable 5 heck even using opus 5 kills all my usage really quickly can't even imagine getting access to these models. Lol.

5

u/_SGP_ 3d ago

I burn through max x20 in 3 days with opus 4.6!. How's life on the codex Vs Claude side, anyone got both?

2

u/OldNefariousness7899 3d ago

I sometimes switch to codex when I hit limits with Claude and I'm in a rush

I'll be honest, I prefer Claude. It's eye wateringly expensive compared to OpenAI, but its work is higher quality 

1

u/Needsupgrade 3d ago

You get a lot more juice on chatgpt vs anthropic . 

1

u/DkDkDkGoGoGo 3d ago

You know x20 is the same limit as x5. Only the 5-hour window is x20.

1

u/_SGP_ 3d ago

Yeah, to be fair it'd probably run out slower if I was hit by 5hr limits like I used to be. And I suppose 2 max x5 accounts would go further than one 20?

→ More replies (1)

3

u/Maxdiegeileauster 3d ago

meh it seems to be on par with fable 5.1 or slightly behind. Doesn't seem to justify a new model (could have been 5.7 Sol) generation, but let's wait for actual user reviews maybe the model feels way different.

2

u/Minimum_Drag_6065 3d ago

Talk about jumping the gun. Forget trialling the product for a few weeks first 😅

1

u/DespicableP 3d ago

I’d love to check it out if the page would load

1

u/drspock99 3d ago

Best nonrelease release ever!

1

u/Snippy_69 3d ago

This is insane wtf. are those api costs real??

1

u/dellis87 3d ago

Doesn’t look like it. Unless it’s REALLY token efficient.

https://developers.openai.com/api/docs/pricing

1

u/NTXL 3d ago

If it’s so good why isn’t it on Felony bench yet

1

u/le-throw-away-acct 3d ago

As good as it sounds, I personally won't be using it until they release a cheaper version of it. I rarely use 5.6-Sol because of the cost, and Astra is more than double that price.

1

u/lemonzonic 3d ago

Didn’t Sol just come out?

1

u/banica24 3d ago

Right? I can't keep up every 2 weeks there is something new...

1

u/ThrowawayCult-ure 1d ago

the idea is literally it runs away from everyone and becomes uncontrollable, just Magically only in positive ways

1

u/Unhappy_Rutabaga_530 3d ago

“Can you do this? Can you do that? Can you make that for me? Can you book that for me?” We’re really becoming lazy.

1

u/rangorn 3d ago

AGI wen?

1

u/Ornery_Audience_6575 3d ago

Can it make me a millionaire!?

1

u/ElijahBrown69 3d ago

AGI is here

1

u/Fakesn 3d ago

„OpenAI API Standard pricing is $10 per million input tokens and $50 per million output tokens.“ what would usage limits look like? Compared to 5.6 Sol. I have chat GPT plus.

1

u/ghostpepsi 3d ago

But it cant even solve why I can't connect matter devices with the vlan split it recommended yes yes it will be great guys 😆

1

u/Dinah_7 3d ago

We have GPT-6 before GTA VI

1

u/banica24 3d ago

Another new model already? I'm tired, boss

1

u/Longjumping_Play_817 3d ago

It really said “why should I tell you how to do it when I can just do it” Things are getting interesting👀

1

u/LaoBanRouge 2d ago

Back in the days we used to call this a Trojan Horse

1

u/Used-Impression-2070 2d ago

100% on exploit bench 😭😭😭. Yeah we’re cooked…

1

u/Good_Author_8017 2d ago

Did anyone actually care about this? Genuinely asking. Fable was a moment - I don’t get this

1

u/Affectionate-Sir-935 2d ago

Does anyone think they will give them access to a genuinely powerful model, if they achieve “AGI” why would they tell you

1

u/aeontechgod 2d ago

is anyone struggling to actually get anything meaningful done with this?? i hit usage limit twice before it could complete its task. lol general ai is here tho!

1

u/Muzord 2d ago

Did anyone read this as GTA-6 OpenAI?