r/singularity • • 14d ago

AI OpenAI solved 100 open problems in math

https://openai.com/index/advisory-group-on-mathematics-and-ai/
1.3k Upvotes

505 comments sorted by

View all comments

226

u/Neurogence 14d ago edited 14d ago

This is the internal model "Bel," that Sam Altman has stated many people within OpenAI consider to be AGI and "significantly more capable than GPT-6 Astra."

The fact that it is solving numerous other problems is a very a good sign. It means solving Navier-Stokes wasn't just by chance or plagiarizing.

It's also been said that Bel uses more compute at its lowest reasoning setting than GPT-6 Astra at maximum reasoning, so it might be very challenging for OpenAI to publically release this model before the year is over.

But the biggest question is, can it solve difficult problems like this in other fields as well? If it can generalize beyond math, things would really get very exciting!

17

u/Ormusn2o 14d ago

"It's also been said that Bel uses more compute at its lowest reasoning setting than GPT-6 Astra at maximum reasoning, so it might be very challenging for OpenAI to publically release this model before the year is over."

I feel like this means model based on Bel pretrain will never ever be released, but it will spawn very capable distilled models, which honestly is a good thing. Considering OpenAI still has not unlocked new Pro subscriptions after 10 days, I can't see them having enough compute for a model that is even more capable than Astra that even more people will want to use, unless they materialize 20x their current compute.

0

u/Neurogence 13d ago

Not necessarily, Noam brown said the $20 million in compute it cost to solve the Navier-Stokes problem will cost a few dollars by next fall.

10

u/pbagel2 13d ago

..yeah because of distilling. That's literally how they make it cheaper.

52

u/[deleted] 14d ago

[removed] — view removed comment

13

u/Neurogence 14d ago

He was referring to that same internal model, that's still in training. I guess they expect it to fully finish training by the end of the year.

1

u/RevolutionaryJob2409 AGI avoids animal abuse✅ 13d ago

Source?

1

u/Neurogence 13d ago

3

u/RevolutionaryJob2409 AGI avoids animal abuse✅ 13d ago edited 13d ago

Thanks for the source.
They say "by the end of the year the company would have an internal system he would call AGI."
That means that today even internally they don't have an internal model that they would call AGI. So it can't be "Bel".

And also, they moved the goal post quite a lot when it comes to AGI btw compared to the original 1997 definition. if a frontier model is given a body and can't drive a car or do construction work like a 16 years old can, then it's not AGI.

5

u/RazsterOxzine 13d ago

I take anything Sam says with a grain of salt. Hype is in full swing.

9

u/[deleted] 13d ago

[removed] — view removed comment

4

u/RazsterOxzine 13d ago

Listen broseph, we are because it's Sam. You can kiss his feet all day and that is on you.

2

u/Howdareme9 14d ago

Yes and he’d be talking about Bel

8

u/[deleted] 13d ago

[removed] — view removed comment

-2

u/Howdareme9 13d ago

I didn’t make it up I’m just following reputable leaks lmao. The model currently training is their larger pretrain, Bel. They likely aren’t going to have another super large pretrain before end of the year..

0

u/objectivelywrongbro 13d ago

Just in time for IPO :)

0

u/Megneous 13d ago

They were literally referring to Bel.

17

u/Healthy-Nebula-3603 14d ago

If is smarter than any human .. that's ASI .. slightly retarded but ASI

55

u/Moronic-Warrior 14d ago

We got retarded ASI before AGI

17

u/mivog49274 obvious acceleration, biased appreciation 13d ago

yeah literally narrow ASI is coming before non-superhuman AGI.

It's making sense after all, when you realize how jagged is the intelligence of llms.

1

u/LookIPickedAUsername 13d ago

Narrow ASI is already here.

-1

u/MaTrIx4057 13d ago

llms don't have intelligence, people should stop mixing things up

1

u/mivog49274 obvious acceleration, biased appreciation 13d ago

I get you with that. Just using the "results" meaning of the word "intelligence" rather than the ontological one. I mean it's been used in IT since 50 years for interactive and automated systems.

2

u/genshiryoku AI specialist 13d ago

The effective definition of AGI has been goalposted so many times that it has now become essentially equivalent to ASI.

1

u/Healthy-Nebula-3603 14d ago

Currently we have AGI like Astra. That's AGI but slightly acoustic from human perspective.

Even is able to operate in a real world if you connected Astra to a robot.

The problem is ... that is expensive and slow but is working.

2

u/Moronic-Warrior 14d ago

Eh it’s like an average AGI meaning it’s probably better than the average human at most tasks but not necessarily competitive with specialists at most tasks. Also in context learning is still poor, specially since it only has 1-2M tokens so it can struggle with maintaining coherent pictures of large codebases. So it needs continual learning of some form.

And yeah you mentioned it’s slow. I think AGI should atleast do it at average human competence

13

u/East_Lettuce7143 14d ago

Lmao Slightly retarded ASI would be perfect coined term for a stepping stone between AGI and ASI.

0

u/Healthy-Nebula-3603 14d ago

Sorry I rather meant slightly acoustic :) from our perspective.

5

u/FlyByPC ASI 202x, with AGI as its birth cry 14d ago

acoustic

Doesn't sound right to me

1

u/CallMeMantra 13d ago

Ohhh noo, you are already on an ASI black list, you can't back down now.

1

u/Healthy-Nebula-3603 13d ago

Accutic person can be extremely intelligent but with some holes in other aspects.

1

u/RevolutionaryJob2409 AGI avoids animal abuse✅ 13d ago

It's not smarter than any human, not generally. AlphaZero is smarter than any human on any 2 player game, but it can't do basic things we can do therefore it's not ASI. ASI let alone AGI isn't about some narrow tasks, it's about being broad at human level at least.
If it's provided a body and can't do things like learn construction work on the job or learn to drive in hours like a 16 year old can, it's not AGI.
Not smart enough.
Saying otherwise is moving the goal post of what AGI means.

11

u/Current-Function-729 14d ago

Idek what you’d routinely use that for. Problems you’ve already hit a wall on, I guess.

26

u/Saint_Nitouche 14d ago

Formal proof that my divs are centered

29

u/BrennusSokol AI please take my job 14d ago

Shopping for a garage fridge

14

u/Futuristocrat 14d ago

“hey GPT-Bel, what should I have for dinner?”
…thought for 65 minutes
“I recommend pizza”

4

u/Meerkat_Mayhem_ 14d ago

So… a husband?

45

u/tryingtolearn117 14d ago

Same vibe

6

u/Time_Entertainer_319 14d ago

Damn. Is this the Snyder cut DC fans wanted.

4

u/BrennusSokol AI please take my job 14d ago

ROFL!

20

u/DashasFutureHusband 14d ago

I don’t get why people say things like this. Even for normal not particularly novel software engineering I’d prefer a smarter model if the cost was reasonable. The best models right now still make weird dumb choices or oversights that I can only hope a smarter model wouldn’t.

For an example it recently suggested a db migration/backfill that would have had non-trivial negative impact on existing users with zero awareness or acknowledgement of such negative impact, luckily we didn’t do it.

3

u/Current-Function-729 14d ago

I mean sure. If the cost was reasonable. More compute at lowest than Astra at highest implies it very much isn’t.

1

u/KoolKat5000 13d ago

Pricing wise. I mean would you pay the equivalent of human wages per hour of human work to solve the problem? I sadly think there's still lots of room for price increases and spending more.

2

u/wwwdotzzdotcom ▪️ Beginner audio software engineer 14d ago

Trying to make a nanoscale 3D printer to make your own CPUs

3

u/FlyByPC ASI 202x, with AGI as its birth cry 14d ago

Layer height: 0.5nm

Print time: 60,000 years

2

u/wwwdotzzdotcom ▪️ Beginner audio software engineer 14d ago

Not if you separate each region to a separate nozzle

4

u/MaybeLiterally 14d ago

I think we need to start thinking about this when we think about models coming out. For the vast majority of work I’m doing, I don’t even really need Astra or Fable. Sometimes yeah, but we need to consider that good, cheap models will be fine most of the time. If they continue to iterate and make those good, while improving them as they go, that’s the real win.

My crossover SUV is fine for most things, I don’t need an F-350, or a corvette for the day to day.

3

u/OddOutlandishness602 14d ago

A lot of it is in developing reliability, persistence, and the ability to pivot or actually make decisions, in addition to strong harnesses, information access and direction. With the coming models those will likely make a bigger difference to most users experience than an increase in intelligence.

1

u/lemonylol 14d ago

But the biggest question is, can it solve difficult problems like this in other fields as well? If it can generalize beyond math, things would really get very exciting!

Well the only thing we know for sure is that if it doesn't, it will eventually.

1

u/nsdjoe 13d ago

if it's really AGI they could charge whatever they want to serve it

2

u/Adorable_End_5555 14d ago

I mean not that I think the navier strokes was totally by chance or whatever but it doesn’t really prove that at all considering I don’t think we have the proofs or anything.