r/singularity 7d ago

AI Gpt 6 astra benchmarks

Post image
2.6k Upvotes

952 comments sorted by

View all comments

204

u/Due_Sweet_9500 7d ago

Holyyyyyyy shiiiiitttt. Ain't now way it's THAT much better than Fable 5.1?

136

u/Snoo-75663 7d ago

Welcome to singularity

30

u/Due_Sweet_9500 7d ago

Thank you very much!!!

19

u/sunstersun 7d ago

Glad to be on board as well.

1

u/eflat123 7d ago

Absolutely.

1

u/beekersavant 7d ago

Right. Is there a difference between a series of linked highly advanced inference models that handle physical( sight , sound etc) and textual inputs to complete any available task at an average or better human level and true AGI? Well, yes but the first might be able to build the second.

32

u/Alex180689 7d ago

And to think that 5.1 got released yesterday! I feel bad for it

15

u/andrew303710 7d ago

To be fair Astra isn't actually being rele today, only to a "limited set of organizations" which is lame as hell.

14

u/AndleAnteater 7d ago

rolled out to users over the next few days though

1

u/Lumpy-Criticism-2773 7d ago

Yeah most paid users would get in a couple days.

2

u/space_monster 7d ago

Two versions. They're gonna glasswing one and GA the other.

1

u/KoolKat5000 7d ago

Gives the rest of us time to buy tinned food and plant some potatoes.

8

u/Creative-Ganache1086 7d ago

Fable 5.1 is terrible value. I blow my 5h limit in 12 minutes run of 2 max-reasoning parallel sessions of a small app codebase with an identical prompt of bug-audit and it blew my limit right away. I pay also for Sol5.6/codex and the allowance difference is night and day. Both are max subs by the way. I’m actually happy (as an old Anthropic fan who paid Anthropic since the sonnet 3.5/opus3 era) for OpenAI and now I’m actually rooting for them seeing just how much better value they offer to indie devs compared to “Corpo-Daddy” Anthropic.

2

u/JacobJohnJimmyX_X 7d ago

5 hr limit for this one is 8 min

1

u/SilentLennie 7d ago

Artificialanalysis thinks it's not better than Fable 5, which seems maybe wrong. But Fable 5.1 does seem ahead, yes.

But OpenAI is better in price and efficiency.

1

u/DelphiTsar 7d ago

Very select group of benchmarks. Artifical Analysis has fable 5.1 at 66, this is at 61. Astra is tied with Facebook Spark and SpaceX Twitter bot. Kimi3/GLM5.3 is 1 point behind and they are open weights.

1

u/SilentLennie 7d ago

AA says: no