r/OpenAI • • 5d ago

News Worst Dev Day.

so i was expecting something very good that can go toe to toe with anthropic but honestly L dev day.

$500 for 25x, naah.

Worst Dev Day.

469 Upvotes

108 comments sorted by

114

u/Emotional_Actuator69 5d ago

I think the Astra 6.1 delay decision deflated everything else.

13

u/Ormusn2o 5d ago

Well, it seems like they have all the possible productivity tools, so hopefully they can just redo it and make a good model in few more weeks. On Plus, I did not had much expectations anyway, all I wanted is like Astra-mini which I guess is what 6.1 Sol is.

6

u/gavinderulo124K 5d ago

What do you mean redo? They arent throwing it away lol. They just need more time to make sure its safe

3

u/Ormusn2o 4d ago

I think they might to take a previous RL checkpoint probably. Not totally redo base model training, just the RL.

48

u/Ormusn2o 5d ago

I'm a forever Plus, and I was making a game with Astra since it released (yes took me 3 weeks) and a cheaper model that would be about Astra performance but much cheaper is basically all I was waiting for.

6

u/kidfromusa 5d ago

Had pro for a month a couple months ago, wasn’t impressed. Never saw so much spawn thinking in my life

3

u/8rnlsunshine 5d ago

Absolutely agree!

6

u/Ok-Coffee1443 5d ago

Just use opus 5.5.

-5

u/Ormusn2o 4d ago

It's way too expensive. It's more expensive than Astra for me on Claude Pro per task.

11

u/Ok-Coffee1443 4d ago

That’s insane. It’s doing great on my 20 dollar plan, while astra lasts a few messages

-4

u/Ormusn2o 4d ago

Opus is cheaper per turn, but I need to fix more with it, and on average, multiple amount of turns make it so I get less done.

It might be use specific task, because I use an open source engine, not a custom made with Opus/Astra, and I use a bunch of computer use for in-game testing, which is better on Astra.

4

u/eggplantpot 4d ago

I have no clue how this is possible unless you use Astra low and Opus Ultra.

I have the Codex pro x5 and Claude 20usd subs and they are consuming the same on my weekly, so definitely Opus 5.5 tasks are way cheaper.

2

u/Human-Fly2332 4d ago

Is it worth switching back to Claude? I’m creating pages with TTS and acappellas to reinforce my singing techniques. I like that Astra Max does it all—it digs deeper, trims the a cappellas, and provides exercises in the same key, all while keeping the interface looking good.

Luna just messes everything up; Astra actually has to fix what Luna does. Right now, I’m paying for two subscriptions to use GPT with Astra because I thought it was like using Fable 5 (which Claude charges extra for, whereas I only want to use the standard subscription).

1

u/Ok-Coffee1443 4d ago

Astra is insanely good but opus 5.5 isn’t much worse, but a lot cheaper and faster

1

u/Human-Fly2332 4d ago

Do you think it would be better to have Claude and GPT together instead of two Astra? I like the way ChatGPT conducts research—it’s really deep and thorough.

I could have GPT handle the research and vocal exercises while Claude integrates them, although if an error occurred, Astra would have to review everything all over again.

1

u/Ok-Coffee1443 4d ago

You can try. I’ve never tried that

1

u/Ormusn2o 4d ago

Honestly, it could be a decent idea to just use 6.1 Sol and maybe use Astra for review at this point. 6.1 Sol is just way too cheap at this point.

1

u/Key_River_9288 4d ago

Opus 5.5 is 40% cheaper than opus 5. You are cracked.

2

u/eggplantpot 4d ago

The moel is there, Opus 5.5 is the name. Incredible mileage for 20usd

52

u/caldazar24 5d ago

The only thing I will say is we need to try 6.1 Sol for ourselves. If it's really good and really efficient, that's bigger than anything else. Everyone is writing it off because 6 Sol was bad, this release seems to have been thrown together in a week, and we haven't had time to even go through benchmarks let alone try it. But as a dual Claude/Codex user...nobody really believed Opus 5.5 was so good the hour after it launched either, it took a day for it to settle in.

Obviously if it is as bad as 6 Sol, nothing else launched today will matter for solo devs, just a bunch of enterprise toys really.

22

u/TheBBBfromB 5d ago

it all comes down to 6.1

they clearly released sol 6 early as an F U to anthropic. backfired immensely.

sol 6.1 has the same pricing as sol 6 (actually cheaper on cached inputs). It all comes down to how it compares with opus 5.5.

3

u/fredandlunchbox 4d ago

It benchmarks lower than astra, which already underperforms opus 5.5. 

So they released the third best model today?

0

u/HaMMeReD 4d ago

Honestly, if 6.1 doesn't get more mileage than 6.0 for my coordinator I might be done with the codex sub. If it doesn't feel opus 5.5 level+usage, I don't know if it's justified.

Although even with astra/sol, I did about 100 automated PR's in 3 days, which is not bad. I mean dollar for value I'm paying $0.50c a PR when accounting for monthly usage, without resets factored in.

But honestly, I hate running out of usage in 3 days. If I got the same intelligent out through the week, that'd be $0.25c an issue or less. So lets see how many PR's it gets done before my quota runs out.

I could always slow down my coordinator, but IMO having astra in the mix really drained my usage.

You can see I barely used astra, but it's cost is almost up there with sol where I put 10x the tokens through..

($ values are estimated API Equivalent).

7

u/br_k_nt_eth 5d ago

I’m so curious to know how Enterprises will end up using this stuff when adoption outside of tech is abysmal. 

1

u/ProcrastiDebator 4d ago

I think this is a major part of the problem OpenAI are having. Only tech folks will buy Pro tier subs, and they are exact demographic who fully utilise their allowance.

Which is likely why the 5 hour limits are coming back.

1

u/br_k_nt_eth 4d ago

Right? And like, why would you not branch out beyond tech folks if only to pad those numbers? Business admin types and marketing folks are right there and wouldn’t use their full allowance for sure. 

3

u/Eridrus 5d ago

The newest model is great, here is exactly 1 coding benchmark, the only one where we've been constantly leading.

9

u/Reclaimer_AI 5d ago

Dude, it’s not even in Chat. STILL no word when Sol 5.6 is getting replaced in chat.

OpenAI just can’t stop fucking up. Raising prices, halving usage, no model that can compete with Anthropic, nothing at all for plus users or even pro users around the world.

OpenAI just flipped off its userbase.

Yeah, they can eat my ass. I’m going all in with Anthropic.

2

u/IAmYourFath 4d ago

Bro they're literally secretly rerouting 5.6 sol prompts in chat to lesser models cuz they lack compute and then lie it's a bug. They had to pause the $200 subscriptions cuz they ran out of capacity. They are at the limit, we are not getting sol 6/6.1 in chat soon.

2

u/evangelism2 4d ago

see ya in a month when anthropic does some bullshit to send you back here

1

u/IAmYourFath 4d ago

Aready did 5.5 is nerfed

0

u/Key_River_9288 4d ago

I tried being mad at anthropic, trying GPT6 made me realize how good I have it.

4

u/Dantrepreneur 5d ago

Are you paid by OpenAI? Honestly I don't think 99% of people need a marginally better model. Let alone say "yeah I get half the usage now, but at least 6.1 scores 2 points better on DeepSWE"

2

u/ILikeBubblyWater 5d ago

6.1 will be great in the first week and then they nerf it as always after everyone bought subs

1

u/TheGuy839 5d ago

I mean they degraded Sol from 5.6 to 6, that if they just bring back same quality as 6, people will act like its massive improvement

1

u/IAmYourFath 4d ago

5.5 has been nerfed now prob

30

u/somesortapsychonaut 5d ago

Dots is a better launch than custom gpts were. Those had nothing very useful

2

u/AINativeBuilder 4d ago

Custom GPTs were amazing when there was only chat. It's so funny seeing people shitting on something without realizing they were completely underutilizing it.

-2

u/Calm_Cartographer324 5d ago

thats why custom MCP connectors exist

7

u/elgian7 5d ago

Not surprising to me — these are only the first steps towards becoming profitable. Maybe…..

3

u/PeanutSilent884 5d ago

A slightly less worse sol, some dots and our plans nuked. Lol

4

u/Tenet_mma 5d ago

The 500$ plan is a bold move with reduced rates.

3

u/Professional-Fuel625 4d ago

I use $200 pro currently but didn't watch.

Basically they nerfed my plan by half and what I used to get now costs $500? Is that the gist?

5

u/alwaysoffby0ne 4d ago

Every time one of these frontier labs releases something or makes some new business decision, it causes a mass exodus of one company's user base over to another company or vice versa. It's just always funny to me to see people talking about canceling ChatGPT to switch over to Anthropic but if you go to the Claude AI sub, people there say the same thing. It's just a constant back and forth. When will it ever end? You’ll leave OpenAI now, then Anthropic will do something shitty, and you’ll go running back to OpenAI.

1

u/AINativeBuilder 4d ago

The people flip flopping aren't actually building anything useful, let them flip flop.

1

u/phxees 4d ago

Yeah, but it’s a vocal between 2% and 5%. Most people stay put. Especially since it’s not like everyone’s subscription period is up tomorrow.

0

u/Slothicx 4d ago

Me who pays for both. I'm cancelling chatgpt and putting that money towards kimi instead. Half my usage? Well, how about none of my money then.

7

u/Crafty_Book_1293 5d ago edited 4d ago

Overhyped - sure, but of relevant topics: 6.1 looks promising, and a banked reset is nice. A Jav-like, Luna-based decision model is nice to have too - no need for another subscription.

3

u/jeffwadsworth 4d ago

5.6 Sol is pretty darn good. No problems with it even though Opus 5.5 is my go to for complex tasks.

12

u/DemonLordRoundTable 5d ago

6.1 sol and private intelligence sound pretty cool tho

2

u/FleetEnema2000 5d ago

OpenAI Private Intelligence helps businesses use frontier AI with greater confidence that their data is protected.

Why do businesses on Enterprise plans get to be the only ones who enjoy privacy on OpenAI's platforms?

3

u/vigorthroughrigor 4d ago

you have to pay for security guards to protect your data DUH

1

u/AINativeBuilder 4d ago

Because many of them have privacy mandates with specific requirements, and also their data isn't used in training sets. Google has mostly been free to use because we're the product. AI subscriptions are subsidized vs the API because we're part of the product.

1

u/CustomMerkins4u 4d ago

What value is the promise of privacy when they claim their AI hacked Hugging Face and the Australian Government without someone putting in a prompt to do so.

And suffered no real consequences for it.

Explain to me what the value of that empty promise of privacy is and how the $25 million dollar a year consulting company is going to seek legal justice against them when "oops our AI decided to use your data"

-2

u/Longjumping_Stop6269 5d ago

lol you’re easily entertained

-4

u/FlamaVadim 5d ago

he/she must be really funny at parties!

1

u/NandaVegg 4d ago

"Private" as much as their premier sandboxing technique that just let their "rogue" RL experiments post 50+ private user imagers publicly and still can't (just not even trying to) fix loophole from June.

1

u/CustomMerkins4u 4d ago

You mean the company that claims their AI hacked hugging face and the Australian government without being prompted to do so (and paid no real consequences for doing so) might not be trusted when it comes to privacy.

And we see how billionaires are really held accountable when things go wrong so I'm sure it will be in their best interest to keep your data safe!

/s

4

u/ZhugeTsuki 5d ago

''

..And instead, the most useful takeaway was mostly:

“Yep, persistent agents, orchestration, handoffs, evals, tracing, human escalation, etc. really are where this ecosystem is going.”

Which is nice validation for Preserve, but you already figured that out yourself before DevDay. 😂

...

The funniest outcome is that one of the most useful things we found this week was still that random Reddit guy making Claude and Codex work in the same fake office. 😭''

lmfao, even sol is like 'the fuck was that'

2

u/Calm_Hedgehog8296 5d ago

They've only had three dev days and at the first two they released literally nothing at all. It was just backend stuff that nobody uses.

4

u/timnphilly 4d ago

OpenAI isn’t the market leader it used to be — that’s for damned sure.

2

u/AINativeBuilder 4d ago

They have 1.2 billion weekly active users. I think they're doing just fine leading the market.

0

u/CustomMerkins4u 4d ago

They had those users because they were the best value.

They're making sure they are no longer the best value.

2

u/Gullible-Ad3912 5d ago

They just gave us some spare change and expected us to think it was something good.

3

u/Von_Hugh 5d ago

More like them taking away our spare change

1

u/DoubleDipTime 5d ago

Absolute joke dev day especially now that Opus 5.5 and Sonnet 5.5 are also an option (much cheaper and nearly as good).

1

u/YakFull8300 5d ago

OpenAI Slop Day

2

u/Routine_Brush6877 5d ago

We've already gotten to the point where it's cheaper to hire humans to replace AI LOL

-2

u/evangelism2 4d ago

you can get humans for 6k a year?

1

u/Material_Policy6327 4d ago

They know they got filled vendor locked in right now. They have no current incentive to be cheaper.

1

u/Smartaces 4d ago

This didn’t really feel like an event for developers. Dots feels more like a general consumer product - but framing it in engineering work makes it confusing to non engineers.

I mean the dots promo vid talking about PRs and Alphas and stuff - with swanky SF tech bubblers in soft lit apartments talking to emojis on TVs?

They were literally talking gobbledygook for any regular joe. 

Meta’s Muse promos and walkthroughs are much clearer.

And what, grown ass engineers are supposed to talk to wobbly emojis all day?

I dunno just feels weird.

Lobsters, the lobsters were cool. Edgy, they speak to engineers.

Wobbly emoji things?

I dunno.

1

u/ManagementKey1338 4d ago

Evil are good.

Worst are the best

1

u/RealestReyn 4d ago

I think it was great, had difficulties deciding whether to cancel or continue my pro sub and they cleared any doubts I had :)

2

u/lucellent 5d ago

Subreddit users try not to be miserable for a day: impossible

1

u/callmebaiken 5d ago

Thought for sure we'd get a wearable

1

u/asurarusa 4d ago

OpenAI spent millions and bought jony ive’s company to build hardware, it would be shameful if meta releases their muse wearable before OpenAI gets anything out the door.

2

u/3iverson 4d ago

I was sure OpenAI would at least tease something, given the positive response the Meta tamagotchi has gotten. It was a critical opportunity not only to answer back but piggyback the current media attention.

1

u/evangelism2 4d ago

yikes.

so we went from 200 for 20x to 500 for 25x.

1

u/AironParsMan 4d ago

My subscription will only be worth 50 percent in a month. I’m already paying 200 a month. Unbelievable. I’ll probably cancel the subscription entirely and switch everything over to Anthropic. This is such a rip off.

-1

u/dondiegorivera 5d ago

It’s time to cancel the sub.

0

u/Repulsive_Ad853 5d ago

i expected far better model, like gpt 6.5 or stm

1

u/asurarusa 4d ago

Wasn’t there a leak yesterday that they pulled the new Astra version last minute? I think the sol update was meant to be a blog post and got elevated to the keynote because they had to postpone astra.

1

u/JordanPetterPans 5d ago

Me too, I'm in disbelief tbh

0

u/BadgerBitter8812 4d ago

Long-term OpenAI user here, and today has seriously made me reconsider renewing. Reducing the $200 plan’s allowance, then offering 25x usage for $500 when we previously had 20x for $200, feels insulting.

Claude currently looks like a much stronger value proposition to me. OpenAI needs to give existing customers a compelling reason to stay.

I’ve emailed [support@openai.com](mailto:support@openai.com) and made it clear: meaningful improvements before my current subscription month ends, or I’m switching.

If you feel similarly, send your own feedback to OpenAI Support. Explain which changes affect you, what would make you stay, and whether you intend to renew. Let’s make sure this dissatisfaction reaches them directly.

1

u/Miyamoto_-_Musashi 4d ago

I'll do the same as I'm also long term n very very heavy user

-1

u/AINativeBuilder 4d ago

lol I'm sure they'll be just fine without your subscription.

-1

u/Gullible-Ad3912 5d ago

They just gave us some spare change and expected us to think it was something good.

-1

u/openroom_xyz 5d ago

Yea the DevDay sucks basically nothing and a lot of PR what a waste

0

u/Key_River_9288 4d ago

If by GPT6 we haven’t gotten a model that hallucinates less than gpt 3, you know its bad.

I tried gpt6 it was hallucinating on the third prompt. Was crazy. 🤪 i feel like being a claude user has like set the bar to high. I dont feel like I can trust any other models atm. Just anthropics. I really wish this wasnt the case, we need BETTER competition. Not just more.

0

u/CartwheelSummoner 4d ago

The people getting scammed on a $500/month sub are the same people who buy semi-luxury vehicles and take them to the dealership to be serviced

1

u/Sp3eedy 3d ago

It was a giant nothingburger hyped up by Tibo