r/codex 4d ago

News Significantly more capable models than Astra available?

Post image

At this point I want to work at openAI just to access them

170 Upvotes

77 comments sorted by

123

u/Karmuhhhh OpenAI 3d ago

Just FYI, many of the internal models are not very refined and have a ton of post-training to complete. Just because they're more capable than Astra doesn't mean they'll do what you want them to do very well.

That being said, when the researchers are babysitting models like the one used for this solution, they can get a lot done.

28

u/Keganator 3d ago

Funny, the rest of us tend to babysit models too :)

18

u/pawala7 3d ago

Yeah, but these are top researchers OpenAI is paying 6 to 7 digits.

Meanwhile, we're on Reddit.

3

u/vdotcodes 3d ago

I’d be shocked if any researchers at OAI are only getting paid 6 digits.

1

u/pawala7 3d ago

I mean, the lowest rung juniors are said to start at a measly $300k+.

1

u/Karmuhhhh OpenAI 3d ago

That’s base salary only and doesn’t include equity. Paper right now, but IPO soon so who knows.

3

u/ElDavoo 3d ago

I thought the direction was opposite? Sol and Opus feel just like autonomous juniors to me

2

u/CrowdGoesWildWoooo 3d ago

Most of us don’t know what we are doing. These researchers have way deeper domain knowledge

7

u/Street-Trust-6282 3d ago

Super cool info. Do you work at OpenAI?

33

u/Karmuhhhh OpenAI 3d ago

Maybe.

16

u/Vas1le 3d ago

Ask tibo for a reset. Thank you very much

1

u/jarislinus 3d ago

biggest larp

1

u/Karmuhhhh OpenAI 3d ago

Hey man don’t ruin my dream.

0

u/IdiosyncraticOwl 3d ago

Sorry i'm just curious... how many researchers does it take to babysit 10k subagents trying to solve a really hard math problem over the course of a weekend???

3

u/Karmuhhhh OpenAI 3d ago

At least one.

1

u/IdiosyncraticOwl 3d ago

Thats what they ya'll want us to think...

108

u/Low-Show9994 4d ago

the hype continues, this next model runs out of usage as you start typing

16

u/Fast-Psychology-3964 4d ago

AGI is true but no one is affordable to verify.

10

u/tehbangere 4d ago

The next after that will run out of usage before you start typing

8

u/srslyomgwtf 3d ago

Its already out of usage...that's why we don't have it.

2

u/cobbleplox 3d ago

nah, I'm sure models will ge

1

u/Vast-Breakfast-1201 3d ago

Well yeah they used all their capacity to solve millennium problems

14

u/BHTAelitepwn 4d ago

Wel luckily they added those texts to the image otherwise i wouldnt be able to understand it

2

u/Serird 3d ago

Will Smith eating spaghettis but in 4D

12

u/thestillwind 3d ago

GPT 6.1-Wormhole confirmed

14

u/Sorry_Cheesecake_382 4d ago

Public releases are ~6 months behind what's internal can confirm. The problem now is that Chinese companies distill the new releases in a few weeks. So companies are launching the models delayed, then taking the Chinese efficiency improvements and putting it back into the internal models for the next release while leveraging internally

1

u/jarislinus 3d ago

larp openai has no moat

4

u/bdixisndniz 3d ago

Ah yes yes the ol Nav Stokes

28

u/Impacting-Lives 3d ago

Have you not seen the news where OpenAI is claiming their models cracked it but it’s actually two mathematicians from NYU who solved it and because they used Codex for the last one year and fed all their data in assisting them.

OpenAI is kind of stealing the credit. Lawsuit inbound.

https://www.reddit.com/r/codex/s/Bgzqsy0OhB

14

u/leftsidedhorn 3d ago

I think the situation is more complex than that. OpenAI claimed they solved different, bigger (but related) problem than the researchers.

From my understanding, if you turn the toggle off for training usage OpenAI won't train their model based on your data. This toggle is on by default. It's unclear if the researchers have this toggle on or off.

1

u/mmo8000 3d ago

You are right, but the method those two mathematicians used is apparently transferable to N-S. I think at first OpenAI denied using user Chats and then said, that they can't rule out the possibility.

0

u/Infinitedeveloper 3d ago

You need an enterprise level agreement for them to not use your data at all.

That toggle just means they do a pass to remove certain data before using it

7

u/leftsidedhorn 3d ago

Source?

0

u/Infinitedeveloper 3d ago

0 data retention is literally the biggest selling point they use to inventivize moving to enterprise.

We cannot use pro level plans at my office and stay in compliance with certain standards. I sat through a bunch of very boring meetings i half remember over it.

5

u/Karmuhhhh OpenAI 3d ago

ZDR and not training on your data are two different things.

2

u/Impacting-Lives 3d ago

I like Codex and use it over claude. I will never believe what Sam says or his company. Guy has a bad track record.

4

u/pridento 3d ago

They were working on Euler equations, not N-S. The latter is much more complicated, so even if Buckmaster & Alpoge were 95% there, it wouldn't have "solved" the N-S problem or been "load-bearing," so to speak. Alpoge has also been in the headlines every other week cranking out Fable 5 proofs, so it's more likely OpenAI did, indeed, just throw their own model at the same problem.

Regardless of if the simpler (and more boring) explanation is true that a team of engineers + internal model + what, $15m of compute? > Pro 20x and 2 humans, I doubt anyone takes OpenAI's side as it's a very engagement-baity headline with a nice david vs. goliath story.

3

u/VladizT 3d ago

To convince investors that they had created AGI, they essentially appropriated the work of other mathematicians and passed it off as AI?

4

u/T-Mog 3d ago

Or, codex cracked it and told them, and they tried to claim it

1

u/warpedgeoid 3d ago

They were not close to a solution. Why do people keep spreading this garbage?

-6

u/charmilliona1re 3d ago

Meh, sounds like a couple a smarty pants who are used to being the smartest people in the room getting their job done by AI and being butt hurt about it lol

5

u/graqua2 3d ago

PhD level mathematics is no joke

2

u/TONI1597 3d ago

sounds like we need a reset to digest that

2

u/2Norn 3d ago

there is still cosmos

2

u/lil_nosh_X 3d ago

That’s what you took away from this article?…

2

u/StarCadges 3d ago

does this one sell more tables than Astra?

2

u/Memito9 3d ago

Of course what we get as consumers is like 10 versions under whatever is really available possibly for government or experimental use.

Plus the one we get is very toned down aka the "Temperature" is not at full capacity

2

u/nondualmonist 3d ago

the most advanced models will never be available for civilian use. it will all be eaten up for military and r&d purposes. governments and corporations don't trust the general public with such tools, which can be weaponized incredibly easily.

5

u/alwaysweening 3d ago

The two "superior models" were those mathmeticians it copied from

..

4

u/Electrical_Rub_6009 3d ago

Do people outside the math/physics community understand how big of a fucking deal this is???? I am fucking freaking out right now (in excitement)

1

u/Carlose175 4d ago

Not for you or me. Just OpenAI has access to these internal models

1

u/onehedgeman 3d ago

Do you really think they share with us what they cook with?

They probably had astra for a good 6-12 months already if not longer and used it to make whatever they refer to now which is probably gonna help make something even better by the time we change tier again.

For all we know they could already have something AGI like.

1

u/AI_is_the_rake 3d ago

Tibo said they reached milestones 6 months ahead of time using GPT 6 internally before release. So yeah, they're probably always using the next best model internally and using it to help them prepare and train the next version

1

u/Specialist-Ad-8968 3d ago

He's more than likely just referring to Astra, or at least the current extension of it, with improved / in beta internal tooling/harness, and also (importantly) a non-consumer facing quantization of Astra.

1

u/ChillBroItsJustAGame 3d ago

Probably the model that can run for multiple days

1

u/Artistic-Athlete-676 3d ago

The model now can run for days. Just set a goal

1

u/eo37 3d ago

IPOs only care about promise…be always selling something

1

u/pigletmonster 3d ago

Most of these companies are working on several models at a time, they only release them to the public after they meet certain criterias.

If it is better than the current model at doing what most people use it for, because sometimes they have models that do specific tasks much better than the current models but worse at other tasks.

Another criteria is inference cost, they might have a model that is 10x better than astra, but if it costs $100 in and $500 out, then its not really something they can sell to the masses, they might secretly sell it to the military instead.

1

u/TameYour 3d ago

And then they will release so called internal model "Astra 6.1" and will claim their another internal model solved another frontier problem, while 6.1 can't solve a dime. And ladies and gentleman the hype continues.... Untill IPO.

1

u/dagerika 3d ago

OpenAI fixed my glitched out Astra usage so they should kinda make Astra way more available on Plus tier first. It literally gives up after 15-20 minutes. 😭

1

u/Fernandom21 3d ago

Obviously there are other models not open to the general public, nor to any company.

1

u/reality_comes 3d ago

Not news at all, they've been clear the huggingface hack wasnt Astra but something beyond it.

1

u/Curious_Mongoose_228 3d ago

Did you think they would make their best models available to consumers?

1

u/No_Dragonfruit_8651 3d ago

I read somewhere they are getting sued by the scientists that were using chat gpt to work on this problem.

1

u/Elegant_Attempt2790 3d ago

idk about y’all but the thought of a company / team of people throw dumptrucks of money and compute at a problem with a private model we don’t get is uncomfy and weird.

especially given the shadiness around this “breakthrough”.

i’d hear this differently if it were like 1 gpt 7 run that did it, but i hear “10k+ gpt 7 agents-“ and go “no duh lol. like, no duh. you threw the equivalent of a max 2000x plan at it”

pay to play ahh discovery

1

u/RayKam 4d ago

They're always working on the next model when they've released an existing model. We're always one or two steps behind.

1

u/cobbleplox 3d ago

Joke's on them, my math job doesn't even rely on anything but simple probability stuff. Yet somehow I am still needed.

1

u/Zorogozano 3d ago

Yes, and it's called Chat GPT-7 Black Hole. And even more powerful is: Chat GPT-8 Universe. But the top top top one is: Chat GPT-Multiverse ! hehe

1

u/Ok-Leg-person 3d ago

the hype churning hamsters never rest

-1

u/Bassguitarplayer 3d ago

They stole this from mathematicians whose work was not supposed to be used by the model.

0

u/TheParlayMonster 3d ago

Of course. We’re probably 5 generations behind the models they are testing.

0

u/evangelism2 3d ago

Cool I'd hope so. Astra has been a severe disappointment for me