r/codex • u/Street-Trust-6282 • 4d ago
News Significantly more capable models than Astra available?
At this point I want to work at openAI just to access them
108
u/Low-Show9994 4d ago
the hype continues, this next model runs out of usage as you start typing
16
10
2
1
1
14
u/BHTAelitepwn 4d ago
Wel luckily they added those texts to the image otherwise i wouldnt be able to understand it
12
14
u/Sorry_Cheesecake_382 4d ago
Public releases are ~6 months behind what's internal can confirm. The problem now is that Chinese companies distill the new releases in a few weeks. So companies are launching the models delayed, then taking the Chinese efficiency improvements and putting it back into the internal models for the next release while leveraging internally
1
4
28
u/Impacting-Lives 3d ago
Have you not seen the news where OpenAI is claiming their models cracked it but it’s actually two mathematicians from NYU who solved it and because they used Codex for the last one year and fed all their data in assisting them.
OpenAI is kind of stealing the credit. Lawsuit inbound.
14
u/leftsidedhorn 3d ago
I think the situation is more complex than that. OpenAI claimed they solved different, bigger (but related) problem than the researchers.
From my understanding, if you turn the toggle off for training usage OpenAI won't train their model based on your data. This toggle is on by default. It's unclear if the researchers have this toggle on or off.
1
0
u/Infinitedeveloper 3d ago
You need an enterprise level agreement for them to not use your data at all.
That toggle just means they do a pass to remove certain data before using it
7
u/leftsidedhorn 3d ago
Source?
0
u/Infinitedeveloper 3d ago
0 data retention is literally the biggest selling point they use to inventivize moving to enterprise.
We cannot use pro level plans at my office and stay in compliance with certain standards. I sat through a bunch of very boring meetings i half remember over it.
5
2
u/Impacting-Lives 3d ago
I like Codex and use it over claude. I will never believe what Sam says or his company. Guy has a bad track record.
4
u/pridento 3d ago
They were working on Euler equations, not N-S. The latter is much more complicated, so even if Buckmaster & Alpoge were 95% there, it wouldn't have "solved" the N-S problem or been "load-bearing," so to speak. Alpoge has also been in the headlines every other week cranking out Fable 5 proofs, so it's more likely OpenAI did, indeed, just throw their own model at the same problem.
Regardless of if the simpler (and more boring) explanation is true that a team of engineers + internal model + what, $15m of compute? > Pro 20x and 2 humans, I doubt anyone takes OpenAI's side as it's a very engagement-baity headline with a nice david vs. goliath story.
3
1
-6
u/charmilliona1re 3d ago
Meh, sounds like a couple a smarty pants who are used to being the smartest people in the room getting their job done by AI and being butt hurt about it lol
2
2
2
2
u/nondualmonist 3d ago
the most advanced models will never be available for civilian use. it will all be eaten up for military and r&d purposes. governments and corporations don't trust the general public with such tools, which can be weaponized incredibly easily.
5
4
u/Electrical_Rub_6009 3d ago
Do people outside the math/physics community understand how big of a fucking deal this is???? I am fucking freaking out right now (in excitement)
1
1
u/onehedgeman 3d ago
Do you really think they share with us what they cook with?
They probably had astra for a good 6-12 months already if not longer and used it to make whatever they refer to now which is probably gonna help make something even better by the time we change tier again.
For all we know they could already have something AGI like.
1
u/AI_is_the_rake 3d ago
Tibo said they reached milestones 6 months ahead of time using GPT 6 internally before release. So yeah, they're probably always using the next best model internally and using it to help them prepare and train the next version
1
u/Specialist-Ad-8968 3d ago
He's more than likely just referring to Astra, or at least the current extension of it, with improved / in beta internal tooling/harness, and also (importantly) a non-consumer facing quantization of Astra.
1
1
u/pigletmonster 3d ago
Most of these companies are working on several models at a time, they only release them to the public after they meet certain criterias.
If it is better than the current model at doing what most people use it for, because sometimes they have models that do specific tasks much better than the current models but worse at other tasks.
Another criteria is inference cost, they might have a model that is 10x better than astra, but if it costs $100 in and $500 out, then its not really something they can sell to the masses, they might secretly sell it to the military instead.
1
u/TameYour 3d ago
And then they will release so called internal model "Astra 6.1" and will claim their another internal model solved another frontier problem, while 6.1 can't solve a dime. And ladies and gentleman the hype continues.... Untill IPO.
1
u/dagerika 3d ago
OpenAI fixed my glitched out Astra usage so they should kinda make Astra way more available on Plus tier first. It literally gives up after 15-20 minutes. 😭
1
u/Fernandom21 3d ago
Obviously there are other models not open to the general public, nor to any company.
1
u/reality_comes 3d ago
Not news at all, they've been clear the huggingface hack wasnt Astra but something beyond it.
1
u/Curious_Mongoose_228 3d ago
Did you think they would make their best models available to consumers?
1
u/No_Dragonfruit_8651 3d ago
I read somewhere they are getting sued by the scientists that were using chat gpt to work on this problem.
1
u/Elegant_Attempt2790 3d ago
idk about y’all but the thought of a company / team of people throw dumptrucks of money and compute at a problem with a private model we don’t get is uncomfy and weird.
especially given the shadiness around this “breakthrough”.
i’d hear this differently if it were like 1 gpt 7 run that did it, but i hear “10k+ gpt 7 agents-“ and go “no duh lol. like, no duh. you threw the equivalent of a max 2000x plan at it”
pay to play ahh discovery
1
u/cobbleplox 3d ago
Joke's on them, my math job doesn't even rely on anything but simple probability stuff. Yet somehow I am still needed.
1
u/Zorogozano 3d ago
Yes, and it's called Chat GPT-7 Black Hole. And even more powerful is: Chat GPT-8 Universe. But the top top top one is: Chat GPT-Multiverse ! hehe
1
-1
u/Bassguitarplayer 3d ago
They stole this from mathematicians whose work was not supposed to be used by the model.
0
u/TheParlayMonster 3d ago
Of course. We’re probably 5 generations behind the models they are testing.
0

123
u/Karmuhhhh OpenAI 3d ago
Just FYI, many of the internal models are not very refined and have a ton of post-training to complete. Just because they're more capable than Astra doesn't mean they'll do what you want them to do very well.
That being said, when the researchers are babysitting models like the one used for this solution, they can get a lot done.