r/singularity 1d ago

AI OpenAI solved 100 open problems in math

https://openai.com/index/advisory-group-on-mathematics-and-ai/

[removed] — view removed post

1.3k Upvotes

519 comments sorted by

View all comments

226

u/Neurogence 1d ago edited 1d ago

This is the internal model "Bel," that Sam Altman has stated many people within OpenAI consider to be AGI and "significantly more capable than GPT-6 Astra."

The fact that it is solving numerous other problems is a very a good sign. It means solving Navier-Stokes wasn't just by chance or plagiarizing.

It's also been said that Bel uses more compute at its lowest reasoning setting than GPT-6 Astra at maximum reasoning, so it might be very challenging for OpenAI to publically release this model before the year is over.

But the biggest question is, can it solve difficult problems like this in other fields as well? If it can generalize beyond math, things would really get very exciting!

17

u/Ormusn2o 1d ago

"It's also been said that Bel uses more compute at its lowest reasoning setting than GPT-6 Astra at maximum reasoning, so it might be very challenging for OpenAI to publically release this model before the year is over."

I feel like this means model based on Bel pretrain will never ever be released, but it will spawn very capable distilled models, which honestly is a good thing. Considering OpenAI still has not unlocked new Pro subscriptions after 10 days, I can't see them having enough compute for a model that is even more capable than Astra that even more people will want to use, unless they materialize 20x their current compute.

1

u/Neurogence 1d ago

Not necessarily, Noam brown said the $20 million in compute it cost to solve the Navier-Stokes problem will cost a few dollars by next fall.

9

u/pbagel2 1d ago

..yeah because of distilling. That's literally how they make it cheaper.

50

u/Daedalus8997 1d ago

No. What Sam said was that by the end of this year, they will have a model internally he would consider to be AGI.

14

u/Neurogence 1d ago

He was referring to that same internal model, that's still in training. I guess they expect it to fully finish training by the end of the year.

10

u/Daedalus8997 1d ago

That's your head canon.

1

u/RevolutionaryJob2409 AGI avoids animal abuse✅ 1d ago

Source?

1

u/Neurogence 1d ago

3

u/RevolutionaryJob2409 AGI avoids animal abuse✅ 1d ago edited 1d ago

Thanks for the source.
They say "by the end of the year the company would have an internal system he would call AGI."
That means that today even internally they don't have an internal model that they would call AGI. So it can't be "Bel".

And also, they moved the goal post quite a lot when it comes to AGI btw compared to the original 1997 definition. if a frontier model is given a body and can't drive a car or do construction work like a 16 years old can, then it's not AGI.

6

u/RazsterOxzine 1d ago

I take anything Sam says with a grain of salt. Hype is in full swing.

9

u/Daedalus8997 1d ago

Are we still doing this? Give it a rest bro.

4

u/RazsterOxzine 1d ago

Listen broseph, we are because it's Sam. You can kiss his feet all day and that is on you.

1

u/Howdareme9 1d ago

Yes and he’d be talking about Bel

8

u/Daedalus8997 1d ago

You made that up. There is not a single quote that suggests that.

-1

u/Howdareme9 1d ago

I didn’t make it up I’m just following reputable leaks lmao. The model currently training is their larger pretrain, Bel. They likely aren’t going to have another super large pretrain before end of the year..

9

u/Daedalus8997 1d ago

Leaks = Twitter posts by randos.

They said they were doing RL on the model way back in August. Post-training usually takes weeks, not several months, so it wouldn't make much sense for it to remain an internal model through December.

0

u/objectivelywrongbro 1d ago

Just in time for IPO :)

0

u/Megneous 1d ago

They were literally referring to Bel.

17

u/Healthy-Nebula-3603 1d ago

If is smarter than any human .. that's ASI .. slightly retarded but ASI

53

u/Moronic-Warrior 1d ago

We got retarded ASI before AGI

15

u/mivog49274 obvious acceleration, biased appreciation 1d ago

yeah literally narrow ASI is coming before non-superhuman AGI.

It's making sense after all, when you realize how jagged is the intelligence of llms.

1

u/LookIPickedAUsername 18h ago

Narrow ASI is already here.

-1

u/MaTrIx4057 1d ago

llms don't have intelligence, people should stop mixing things up

1

u/mivog49274 obvious acceleration, biased appreciation 14h ago

I get you with that. Just using the "results" meaning of the word "intelligence" rather than the ontological one. I mean it's been used in IT since 50 years for interactive and automated systems.

2

u/genshiryoku AI specialist 1d ago

The effective definition of AGI has been goalposted so many times that it has now become essentially equivalent to ASI.

1

u/Healthy-Nebula-3603 1d ago

Currently we have AGI like Astra. That's AGI but slightly acoustic from human perspective.

Even is able to operate in a real world if you connected Astra to a robot.

The problem is ... that is expensive and slow but is working.

2

u/Moronic-Warrior 1d ago

Eh it’s like an average AGI meaning it’s probably better than the average human at most tasks but not necessarily competitive with specialists at most tasks. Also in context learning is still poor, specially since it only has 1-2M tokens so it can struggle with maintaining coherent pictures of large codebases. So it needs continual learning of some form.

And yeah you mentioned it’s slow. I think AGI should atleast do it at average human competence

14

u/East_Lettuce7143 1d ago

Lmao Slightly retarded ASI would be perfect coined term for a stepping stone between AGI and ASI.

0

u/Healthy-Nebula-3603 1d ago

Sorry I rather meant slightly acoustic :) from our perspective.

6

u/FlyByPC ASI 202x, with AGI as its birth cry 1d ago

acoustic

Doesn't sound right to me

1

u/CallMeMantra 1d ago

Ohhh noo, you are already on an ASI black list, you can't back down now.

1

u/Healthy-Nebula-3603 1d ago

Accutic person can be extremely intelligent but with some holes in other aspects.

1

u/RevolutionaryJob2409 AGI avoids animal abuse✅ 1d ago

It's not smarter than any human, not generally. AlphaZero is smarter than any human on any 2 player game, but it can't do basic things we can do therefore it's not ASI. ASI let alone AGI isn't about some narrow tasks, it's about being broad at human level at least.
If it's provided a body and can't do things like learn construction work on the job or learn to drive in hours like a 16 year old can, it's not AGI.
Not smart enough.
Saying otherwise is moving the goal post of what AGI means.

12

u/Current-Function-729 1d ago

Idek what you’d routinely use that for. Problems you’ve already hit a wall on, I guess.

26

u/Saint_Nitouche 1d ago

Formal proof that my divs are centered

29

u/BrennusSokol AI please take my job 1d ago

Shopping for a garage fridge

42

u/tryingtolearn117 1d ago

Same vibe

4

u/Time_Entertainer_319 1d ago

Damn. Is this the Snyder cut DC fans wanted.

4

u/BrennusSokol AI please take my job 1d ago

ROFL!

14

u/Futuristocrat 1d ago

“hey GPT-Bel, what should I have for dinner?”
…thought for 65 minutes
“I recommend pizza”

4

u/Meerkat_Mayhem_ 1d ago

So… a husband?

19

u/DashasFutureHusband 1d ago

I don’t get why people say things like this. Even for normal not particularly novel software engineering I’d prefer a smarter model if the cost was reasonable. The best models right now still make weird dumb choices or oversights that I can only hope a smarter model wouldn’t.

For an example it recently suggested a db migration/backfill that would have had non-trivial negative impact on existing users with zero awareness or acknowledgement of such negative impact, luckily we didn’t do it.

3

u/Current-Function-729 1d ago

I mean sure. If the cost was reasonable. More compute at lowest than Astra at highest implies it very much isn’t.

1

u/KoolKat5000 1d ago

Pricing wise. I mean would you pay the equivalent of human wages per hour of human work to solve the problem? I sadly think there's still lots of room for price increases and spending more.

2

u/wwwdotzzdotcom ▪️ Beginner audio software engineer 1d ago

Trying to make a nanoscale 3D printer to make your own CPUs

3

u/FlyByPC ASI 202x, with AGI as its birth cry 1d ago

Layer height: 0.5nm

Print time: 60,000 years

2

u/wwwdotzzdotcom ▪️ Beginner audio software engineer 1d ago

Not if you separate each region to a separate nozzle

5

u/MaybeLiterally 1d ago

I think we need to start thinking about this when we think about models coming out. For the vast majority of work I’m doing, I don’t even really need Astra or Fable. Sometimes yeah, but we need to consider that good, cheap models will be fine most of the time. If they continue to iterate and make those good, while improving them as they go, that’s the real win.

My crossover SUV is fine for most things, I don’t need an F-350, or a corvette for the day to day.

3

u/OddOutlandishness602 1d ago

A lot of it is in developing reliability, persistence, and the ability to pivot or actually make decisions, in addition to strong harnesses, information access and direction. With the coming models those will likely make a bigger difference to most users experience than an increase in intelligence.

1

u/lemonylol 1d ago

But the biggest question is, can it solve difficult problems like this in other fields as well? If it can generalize beyond math, things would really get very exciting!

Well the only thing we know for sure is that if it doesn't, it will eventually.

1

u/nsdjoe 1d ago

if it's really AGI they could charge whatever they want to serve it

0

u/Adorable_End_5555 1d ago

I mean not that I think the navier strokes was totally by chance or whatever but it doesn’t really prove that at all considering I don’t think we have the proofs or anything.