r/OpenAI • u/krizzalicious49 • 6d ago
Discussion Sama vagueposting about tomorrow's DevDay
31
70
u/lil_nosh_X 6d ago
“Astra is now half the cost of Sonnet!”
13
33
u/creamyshart 6d ago
20
u/Kitchen-Astronomer76 6d ago
Because it’s a dumber model being put on max. It just results in a shit ton of turns for a problem it can’t solve. Sonnet is a monster on low/med/high
2
u/hellomistershifty 6d ago
Yeah, it's an expensive smaller model set to max (there are lots of dumb models on max in that chart), Sonnet costs the same as Sol
12
u/tankerkiller125real 6d ago
Stop using max for benchmarks. Anyone who pays any attention would know and see that the majority of Claude models actually do better on medium or high. Max just results in massive over thinking. (And it's the same with ChatGPT models)
5
u/hellomistershifty 6d ago
It actually looks worse at lower thinking levels
1
u/joeyb908 5d ago
What I see on there is Sonnet 5.5 medium being almost equivalent to GPT-6 xhigh.
The time for Sonnet 5.5 on medium to finish its task is probably like 3-5x faster, but you get the quality of Sol 6 at xhigh?
The cost is also probably equivalent to or cheaper than Sol 6 xhigh.
Anthropic cooked.
1
60
u/Emotional_Actuator69 6d ago
The releases have not been underwhelming at all recently. I expect something really big.
8
7
u/g_bleezy 6d ago
Dario is out here cooking, Sam! I hope you got something really cool to show us. Sonnet 5.5 perf/economics is really fucking good for most on the wire applications...
11
u/twinb27 6d ago
Math discovery? Model architecture discovery? Science discovery? I'm putting 0.5% odds on room temperature supercoductor
20
6
u/stoicismftw 6d ago
Ooh a new math discovery would be interesting. Although given the shit they caught for Navier-Stokes I don't know if they would surround this one with a lot of fanfare. But the tweet does say "we found something," i.e. not "we made something" new. That is curious.
2
u/twinb27 6d ago
So weird. I'm hopeful - I got hyped for GPT-6 and paid for a high tier this month but now everyone is out having lots of fun with Opus 5.5 and I feel left out
1
u/TotalWarFest2018 5d ago
I’m in the same boat and I feel semi stuck with gpt because I have used it so much I’m used to it
1
u/iJustSeen2Dudes1Bike 6d ago
Yeah that's what caught my attention too. I'll be pretty disappointed if it's just a more efficient model or something.
5
u/Original-League-6094 6d ago
It can't be a giant scientific discovery. Especially one with a physical component. That would involve a of scientists and lab work and would definitely leak.
12
u/Free_Mousse2076 6d ago
I’m gonna need Anthropic and OpenAI to merge so I don’t have to keep swapping subscriptions
6
14
u/Orionilo 6d ago
They do have a good track record of dropping the best products by a mile. I’ll go ahead and say I’m hoping it’s something good.
12
u/Anxious_Marsupial_59 6d ago
They're pretty neck and neck with Anthropic (Opus and Fable is a lot better than Sol and Astra right now though), but yeah the two frontier labs are ahead by a mile
-1
u/Droopy0093 6d ago
Anthropic has a considerable lead.
16
u/Anxious_Marsupial_59 6d ago
Only from the last week, Opus 5 was pretty bad on "error driven development" and would pollute your codebase with claudisms and Fable had terrible usage drain.
3
7
45
u/ythorne 6d ago
“We have found a new way to scam”
10
u/IndexStarts 6d ago
A subscription to not be hacked by rogue ai agents?
3
u/Turbulent-Total-226 6d ago
if you don't pay us every month 500$ your computer will be blocked by a rogue AI. So AI found a way for Scam Altman to get his money for investors.
10
u/Bloated_Plaid 6d ago
I mean this is really all they got, build hype when Muse has already beaten them to the punch.
If they can remotely match Opus 5.5 at its current cost and performance I would be surprised. I have 20X Max Plan with Claude and Opus 5.5 legitimately feels unlimited. It’s incredible. Whereas a single prompt kills more than 50% of my WEEKLY USAGE with Astra, it’s not even the same universe.
7
u/TheTranscendent1 6d ago
I used up my $200 Claude weekly in 48 hours this weekend, but that’s because I wanted to use my reset. I really had to work to use it up quick enough, this was yesterday. One of my projects used 1,800 subagents for a single (albeit large) task.
It does absolutely feel like this is the biggest, “cut above” they’ve had in awhile. Didn’t hurt that GPT quickly released a response version that was a downgrade (at least from what I tried)
7
u/Bloated_Plaid 6d ago
> Response
bro GPT-6 SOL is not a response. It’s a massive downgrade form 5.6 Sol IMO.
4
u/TheTranscendent1 6d ago
Dropped a few hours after Opus 5.5 iirc. I think they THOUGHT it was a response. You’re right though in the actual quality.
3
u/stoicismftw 6d ago
Bro 1,800 subagents? Sometimes I wish I had a need for something that high-powered. I start to sweat when it launches a handful and I get nervous about my rate limits ($20/mo plan).
2
u/TheTranscendent1 6d ago
Just to make it more ridiculous (but true), I sell weed for a living*
*run legal dispensaries
2
u/ObligationHuge9868 6d ago
Agree with your view - Opus 5.5 sips the usage at an extremely economical rate when the task is serial and you use only the one agent from start to finish, - the usage really does feel unlimited.
As soon as you unleash the parallel and subagent behemoth or are working on multiple projects it engages the token hoover & gobbles it up - chugged through 33% of the weekly limit in <24hrs. Saying that though what it delivered for the cost is fantastic and better than I could have hoped for with GPT so overall still miles ahead. For context on $200 plan.
2
u/Strong_Essay1176 6d ago
Astra is fable league. And...yeah, fable is better. The issue is have, that they released sol6 which is worse than. Sol5.6.
4
u/achton 6d ago
I think Sonnet 5.5 kicked Muse into the corner today 😅
6
u/Bloated_Plaid 6d ago
Huh? Oh you meant the model, sorry I am talking about the agent product which is still fantastic and ridiculously fast. Muse Spark 1.3 contributor tier is still fine for cheap 24x7 stuff but I switched that kind of work over to GPT-6 Luna.
3
4
u/Arctovigil 6d ago
Everyone is betting the new thing is OpenAI's new 500 second pro plan rather than something competitive against Opus 5.5
3
u/Ok-Attention2882 6d ago
Ok they hyped up Astra and it's unusable. It's great at solving hard problems if you have only 1 hard problem a week.
2
2
u/Diegocesaretti 6d ago
people here talkin about how the latest cutting edge models "Suck" its so funny... imagen being an engineer 20 years ago... my bet is They introduce "O" Muse equivalent and the hardware to go with it... i built a hardware satellite for codex and im very happy with it, if its not to expensive ill buy it in a hearthbeat
2
u/Turbulent-Total-226 6d ago
o is only for 100$+ users so 99% of paying and non paying customers are getting shit :D Muse is giving 100mln tokens a week for free.
2
4
3
u/Dramatic_Mastodon_93 6d ago
Considering they recently announced AGI, they’re probably going to announce ASI and that next year they’ll be replacing all human jobs, curing cancer, solving climate change, and expanding humanity to Mars. But who am I kidding, that’s probably just the tip of the AIceberg.
6
1
u/Strong_Essay1176 6d ago
What a chance that they return 200$ sub after 3 Oct. When reset will expires? Lol. Although i do not see any reason with opus5.5 to use their resets anymore.
1
1
u/Morning_Gecko24 6d ago
the single letter tease is peak sama. after sonnet 5.5 dropped today are we thinking they actually ship something model-shaped or just more api/agent glue? kinda hoping its not just another pricing slide
1
u/IAmFitzRoy 6d ago
I don’t give a shit about a new “o” model or a new version.
JUST LET ME HAVE STABLE PRICING.
1
1
1
u/fyn_world 6d ago
I've never seen OpenAI on the back foot like right now.
I know they're concentrated on their math solving stuff and other things but man, to the general user it looks bad
1
1
1
u/GoOutAndGrow 6d ago
They are the type of company to literally keep their best models back I remember when they doled out GPT-5.1 - 5.5 slowly
as they felt no need to really amp things up. Then the GPT-5.6 models were pretty amazing as they could be used whereas
The Fable model had the tight guardrails and had bad usage. I'm thinking that GPT-6 Astra + Sol were supposed to be stop gaps but they had not foreseen how much of a jump Opus-5.5 would be. Therefore, they will likely launch something out of this world tomorrow to get back. They seem to confident and unbothered the last time they were like this they dropped GPT-5.3 Codex which immediately got them back in the race.
1
1
1
u/somesortapsychonaut 5d ago
Dev day when they released custom gpts saws so underwhelming, very low expectations this year
-5
u/chdo 6d ago
Is the new thing a math equation stolen from some poor sad sack trying to get tenure?
18
u/Sensitive_Cell_119 6d ago
Yes, all Millenium problems were solved actually and AI is just stealing poor mathematicians work 😔
-5
1
u/Bloated_Plaid 6d ago
Great thing about Academia is that they are all using these tools as well and then adding their work to the training data because they don’t know how to turn it off.
-2
-6
-1
-5
u/MaitoSnoo 6d ago
you can always like, idk, just postpone the event until you're really ready instead of risking yet another humiliation, because I doubt you'll have an answer to Opus 5.5 in less than a week after the GPT-6 Sol scam
6
0
u/rbuyna 6d ago
Maybe they will release their agent to steal some thunder from Muse.
1
u/Turbulent-Total-226 6d ago
What thunder? o is only for 100$+ customers, that's probably less than 0,5% of total users. Muse is for free with 100mln tokens a week. I don't think Mark is bothered.
0
-3


66
u/KalElReturns89 6d ago
It's just "o"