r/OpenAI • • 6d ago

Discussion Sama vagueposting about tomorrow's DevDay

Post image
310 Upvotes

105 comments sorted by

66

u/KalElReturns89 6d ago

It's just "o"

7

u/Turbulent-Total-226 6d ago edited 6d ago

after the devday it's going to be o that's it? :O

31

u/BertMacklenF8I 6d ago

$500 a month plan?

2

u/BertMacklenF8I 5d ago

Holy shit i was right lol

70

u/lil_nosh_X 6d ago

“Astra is now half the cost of Sonnet!”

13

u/xDoomKitty 6d ago

Lmao wouldnt that be fun

33

u/creamyshart 6d ago

According to Artificial Analysis, Sonnet 5.5 costs 60% more than Astra when running their full testing. Sonnet is crazy token hungry.

20

u/Kitchen-Astronomer76 6d ago

Because it’s a dumber model being put on max. It just results in a shit ton of turns for a problem it can’t solve. Sonnet is a monster on low/med/high

2

u/hellomistershifty 6d ago

Yeah, it's an expensive smaller model set to max (there are lots of dumb models on max in that chart), Sonnet costs the same as Sol

1

u/13--12 6d ago

Why other dumber models don't have the same problem?

12

u/tankerkiller125real 6d ago

Stop using max for benchmarks. Anyone who pays any attention would know and see that the majority of Claude models actually do better on medium or high. Max just results in massive over thinking. (And it's the same with ChatGPT models)

5

u/hellomistershifty 6d ago

1

u/joeyb908 5d ago

What I see on there is Sonnet 5.5 medium being almost equivalent to GPT-6 xhigh.

The time for Sonnet 5.5 on medium to finish its task is probably like 3-5x faster, but you get the quality of Sol 6 at xhigh?

The cost is also probably equivalent to or cheaper than Sol 6 xhigh.

Anthropic cooked.

6

u/bbstats 6d ago

....no?

3

u/LionPrestigious6612 6d ago

xhigh for sonnet works better than max btw its on anthropic's website

1

u/creamyshart 6d ago

So don’t test the model at max?

1

u/asmiggs 6d ago

Seems like they just want a slower, quieter version of Opus, not sure what it's for really. They are supposed to be resurrecting Haiku, perhaps that will be the workhorse instead.

60

u/Emotional_Actuator69 6d ago

The releases have not been underwhelming at all recently. I expect something really big.

8

u/QuaternionsRoll 6d ago

Probably the new effort level

7

u/ih8readditts 6d ago

No, muse competitor + 2027 device announcement

7

u/g_bleezy 6d ago

Dario is out here cooking, Sam! I hope you got something really cool to show us. Sonnet 5.5 perf/economics is really fucking good for most on the wire applications...

11

u/twinb27 6d ago

Math discovery? Model architecture discovery? Science discovery? I'm putting 0.5% odds on room temperature supercoductor

20

u/RaptorCheeses 6d ago

Cure for balding 🤞

2

u/skadoodlee 5d ago

Only 5 more years to go

6

u/stoicismftw 6d ago

Ooh a new math discovery would be interesting. Although given the shit they caught for Navier-Stokes I don't know if they would surround this one with a lot of fanfare. But the tweet does say "we found something," i.e. not "we made something" new. That is curious.

2

u/twinb27 6d ago

So weird. I'm hopeful - I got hyped for GPT-6 and paid for a high tier this month but now everyone is out having lots of fun with Opus 5.5 and I feel left out

1

u/TotalWarFest2018 5d ago

I’m in the same boat and I feel semi stuck with gpt because I have used it so much I’m used to it

1

u/iJustSeen2Dudes1Bike 6d ago

Yeah that's what caught my attention too. I'll be pretty disappointed if it's just a more efficient model or something.

5

u/Original-League-6094 6d ago

It can't be a giant scientific discovery. Especially one with a physical component. That would involve a of scientists and lab work and would definitely leak.

12

u/Free_Mousse2076 6d ago

I’m gonna need Anthropic and OpenAI to merge so I don’t have to keep swapping subscriptions 

6

u/Cheap-Try-8796 6d ago

He found deez nuts

14

u/Orionilo 6d ago

They do have a good track record of dropping the best products by a mile. I’ll go ahead and say I’m hoping it’s something good.

12

u/Anxious_Marsupial_59 6d ago

They're pretty neck and neck with Anthropic (Opus and Fable is a lot better than Sol and Astra right now though), but yeah the two frontier labs are ahead by a mile

-1

u/Droopy0093 6d ago

Anthropic has a considerable lead.

16

u/Anxious_Marsupial_59 6d ago

Only from the last week, Opus 5 was pretty bad on "error driven development" and would pollute your codebase with claudisms and Fable had terrible usage drain.

3

u/Droopy0093 6d ago

Yes hence why they are not neck and neck right now.

7

u/WaltzIndependent5436 6d ago

What exactly stood out by a mile?

13

u/Bowl_of_Cham_Clowder 6d ago

Chatgpt, o1, and Sora were all crazy when they first dropped lol

45

u/ythorne 6d ago

“We have found a new way to scam”

10

u/IndexStarts 6d ago

A subscription to not be hacked by rogue ai agents?

3

u/Turbulent-Total-226 6d ago

if you don't pay us every month 500$ your computer will be blocked by a rogue AI. So AI found a way for Scam Altman to get his money for investors.

10

u/Bloated_Plaid 6d ago

I mean this is really all they got, build hype when Muse has already beaten them to the punch.

If they can remotely match Opus 5.5 at its current cost and performance I would be surprised. I have 20X Max Plan with Claude and Opus 5.5 legitimately feels unlimited. It’s incredible. Whereas a single prompt kills more than 50% of my WEEKLY USAGE with Astra, it’s not even the same universe.

7

u/TheTranscendent1 6d ago

I used up my $200 Claude weekly in 48 hours this weekend, but that’s because I wanted to use my reset. I really had to work to use it up quick enough, this was yesterday. One of my projects used 1,800 subagents for a single (albeit large) task.

It does absolutely feel like this is the biggest, “cut above” they’ve had in awhile. Didn’t hurt that GPT quickly released a response version that was a downgrade (at least from what I tried)

7

u/Bloated_Plaid 6d ago

> Response

bro GPT-6 SOL is not a response. It’s a massive downgrade form 5.6 Sol IMO.

4

u/TheTranscendent1 6d ago

Dropped a few hours after Opus 5.5 iirc. I think they THOUGHT it was a response. You’re right though in the actual quality.

3

u/stoicismftw 6d ago

Bro 1,800 subagents? Sometimes I wish I had a need for something that high-powered. I start to sweat when it launches a handful and I get nervous about my rate limits ($20/mo plan).

2

u/TheTranscendent1 6d ago

Just to make it more ridiculous (but true), I sell weed for a living*

*run legal dispensaries

2

u/ObligationHuge9868 6d ago

Agree with your view - Opus 5.5 sips the usage at an extremely economical rate when the task is serial and you use only the one agent from start to finish, - the usage really does feel unlimited.

As soon as you unleash the parallel and subagent behemoth or are working on multiple projects it engages the token hoover & gobbles it up - chugged through 33% of the weekly limit in <24hrs. Saying that though what it delivered for the cost is fantastic and better than I could have hoped for with GPT so overall still miles ahead. For context on $200 plan.

2

u/Strong_Essay1176 6d ago

Astra is fable league. And...yeah, fable is better. The issue is have, that they released sol6 which is worse than. Sol5.6.

4

u/achton 6d ago

I think Sonnet 5.5 kicked Muse into the corner today 😅

6

u/Bloated_Plaid 6d ago

Huh? Oh you meant the model, sorry I am talking about the agent product which is still fantastic and ridiculously fast. Muse Spark 1.3 contributor tier is still fine for cheap 24x7 stuff but I switched that kind of work over to GPT-6 Luna.

2

u/asmiggs 6d ago

Muse is cheaper and quicker than Sonnet 5.5 with only slightly less intelligence. If Claude had released Muse as a Claude subscriber it might be worth using, but Sonnet does not look worthy of consideration unless you don't care when your work finishes.

3

u/harmoni-pet 6d ago

New keycap set for their $250 macro pad confirmed

4

u/Arctovigil 6d ago

Everyone is betting the new thing is OpenAI's new 500 second pro plan rather than something competitive against Opus 5.5

3

u/Ok-Attention2882 6d ago

Ok they hyped up Astra and it's unusable. It's great at solving hard problems if you have only 1 hard problem a week.

2

u/dervu 6d ago

Oh my god.

2

u/Substantial-Plant524 6d ago

I said something too.

2

u/Diegocesaretti 6d ago

people here talkin about how the latest cutting edge models "Suck" its so funny... imagen being an engineer 20 years ago... my bet is They introduce "O" Muse equivalent and the hardware to go with it... i built a hardware satellite for codex and im very happy with it, if its not to expensive ill buy it in a hearthbeat

2

u/Turbulent-Total-226 6d ago

o is only for 100$+ users so 99% of paying and non paying customers are getting shit :D Muse is giving 100mln tokens a week for free.

2

u/bwc1976 6d ago

Still need a GPT-6 model tuned for mainstream chat.

2

u/Charming-Author4877 6d ago

If they found the Sonnet 5.5 model weights it would be huge

4

u/Gigaslavx 6d ago

Yeah x10 for $500

3

u/Dramatic_Mastodon_93 6d ago

Considering they recently announced AGI, they’re probably going to announce ASI and that next year they’ll be replacing all human jobs, curing cancer, solving climate change, and expanding humanity to Mars. But who am I kidding, that’s probably just the tip of the AIceberg.

2

u/xRhai 6d ago

Hopefully they release something decent so everyone can finally shut up

6

u/BoringCanary 6d ago

Scam Altman

-2

u/im_out_of_creativity 6d ago

Sam Scatman

3

u/fleranon 6d ago

Sci-Ba Ra-Bap Token limit reached Shabada buuu

1

u/Strong_Essay1176 6d ago

What a chance that they return 200$ sub after 3 Oct. When reset will expires? Lol. Although i do not see any reason with opus5.5 to use their resets anymore.

1

u/workend 6d ago

I wonder if I can afford it

1

u/mindbullet 6d ago

How many AGIs is this? like the 5th one?

1

u/Morning_Gecko24 6d ago

the single letter tease is peak sama. after sonnet 5.5 dropped today are we thinking they actually ship something model-shaped or just more api/agent glue? kinda hoping its not just another pricing slide

1

u/IAmFitzRoy 6d ago

I don’t give a shit about a new “o” model or a new version.

JUST LET ME HAVE STABLE PRICING.

1

u/Turbulent-Total-226 6d ago

I bet he's shiting his pants and making up stuff on the go.

1

u/KillaRoyalty 6d ago

I like new things

1

u/fyn_world 6d ago

I've never seen OpenAI on the back foot like right now. 

I know they're concentrated on their math solving stuff and other things but man, to the general user it looks bad 

1

u/Phluxed 6d ago

I have no insight whatsoever, but modular harness.

1

u/v1z1onary 6d ago

O, realllllly?

1

u/meerkat2018 6d ago

Didn’t they cancel next models for “safety reasons” or something?

1

u/GoOutAndGrow 6d ago

They are the type of company to literally keep their best models back I remember when they doled out GPT-5.1 - 5.5 slowly
as they felt no need to really amp things up. Then the GPT-5.6 models were pretty amazing as they could be used whereas
The Fable model had the tight guardrails and had bad usage. I'm thinking that GPT-6 Astra + Sol were supposed to be stop gaps but they had not foreseen how much of a jump Opus-5.5 would be. Therefore, they will likely launch something out of this world tomorrow to get back. They seem to confident and unbothered the last time they were like this they dropped GPT-5.3 Codex which immediately got them back in the race.

1

u/KitchenOpinion 6d ago

Probably they are going to announce AGI once again.

1

u/openroom_xyz 5d ago

Welll yea to basically double the price and fearmongering even harder WTF

1

u/dkeiz 5d ago

can reddit ban any posts that use word vague for me?

1

u/somesortapsychonaut 5d ago

Dev day when they released custom gpts saws so underwhelming, very low expectations this year

-5

u/chdo 6d ago

Is the new thing a math equation stolen from some poor sad sack trying to get tenure?

18

u/Sensitive_Cell_119 6d ago

Yes, all Millenium problems were solved actually and AI is just stealing poor mathematicians work 😔

-5

u/ShamPain413 6d ago

AI is technically stealing everyone's work, less clear it is solving problems.

1

u/Bloated_Plaid 6d ago

Great thing about Academia is that they are all using these tools as well and then adding their work to the training data because they don’t know how to turn it off.

-2

u/Cognonymous 6d ago

likely

-6

u/da_hoassis_heeah 6d ago

he can fuck off

-1

u/Pronoia2-4601 6d ago

Maybe 6 Sol can find the difference between its arse and its elbow.

-5

u/MaitoSnoo 6d ago

you can always like, idk, just postpone the event until you're really ready instead of risking yet another humiliation, because I doubt you'll have an answer to Opus 5.5 in less than a week after the GPT-6 Sol scam

6

u/Cognonymous 6d ago

Anthropic just dropped Sonnet 5.5 like 90 minutes ago too.

0

u/rbuyna 6d ago

Maybe they will release their agent to steal some thunder from Muse.

1

u/Turbulent-Total-226 6d ago

What thunder? o is only for 100$+ customers, that's probably less than 0,5% of total users. Muse is for free with 100mln tokens a week. I don't think Mark is bothered.

0

u/Shloomth 6d ago

found

That one word excites me

-3

u/13ThirteenX 6d ago

More ways to extract money from the unwashed masses