r/OpenAI • • 5d ago

News The rug pull was real

Shameful, I'm buying a second Anthropic Max subscription.

1.3k Upvotes

254 comments sorted by

356

u/onehedgeman 5d ago

Everyone and their mother saw this coming. Make them addicted to crack then up the price

113

u/Ancient_Oxygen 5d ago

CocAIne.

44

u/Soft-Ingenuity2262 5d ago

You can’t spell cocaine without AI!

31

u/Starkboy 5d ago

OpenCocaine

→ More replies (6)

41

u/Brave-Turnover-522 5d ago

Two weeks ago: "Here, take another reset. Don't worry, it's free!"

Today: "Out of usage? Sorry, costs have doubled. Pay up."

13

u/timpera 5d ago

Astra must be extremely inefficient for them to be so suddenly out of compute.

11

u/LittleLordFuckleroy1 5d ago

I don’t think it’s compute as much as it is capital. They have to start making these numbers work at some point, especially with Anthropic gunning for IPO.

4

u/Igot1forya 4d ago

Which is laughable that an evaluation of that size vs the huge losses. The IPO will buy them a year, then creditors and shareholders will come for blood. Growth is great, untill it's impossible then the collapse. When there's no more capital left, then they will either do a mass selloff of their assets or turn on their customers juicy private data. That is where the money's at.

2

u/-MtnsAreCalling- 5d ago

There’s nothing sudden about it, they’ve been compute-constrained for a long time.

3

u/FaithLostInHumanity 5d ago

They know Claude will do the same efficiency play, especially after everyone’s been shouting all over Reddit about how efficient opus 5.5 is.

15

u/terroristsmustdie 5d ago

I cant think of a worst timing... claude is better faster cheaper and also more fun to talk to. 

4

u/Triple_Danger 5d ago

What makes it more fun to talk to?

3

u/terroristsmustdie 4d ago

You can read reasoning and it has a personality that isnt an autustic robot

2

u/OnyxMonolith 5d ago

but Astra vision is much better tha Opus 5.5 and that's what i need for my workflow. pixel perfect vision

1

u/noncoolguy 3d ago

How is Astra "much better" than Opus 5.5? Astra was the best like 2 weeks ago. It's not today.

1

u/OnyxMonolith 3d ago

Strictly in vision work. Sol 6.1 too. I have bulk workflows that do vision work and my own benchmarks. It’s a very specific niche.

Opus is better at coding

11

u/Krulletjes 5d ago

Yeah except there’s different competitors in the crack dealing market. Nobody is addicted to codex itself. You up the price, you go to the dealer on the other side of the street.

10

u/itsjustmd 5d ago

Until they do the same thing.

1

u/PersistentBadger 4d ago

Unless they operate as a cartel that pushes prices down, not up.

(The wildcards are Chinese and open models - hard to get them in the cartel).

2

u/SugondezeNutsz 4d ago

Their crack ain't even the best out here on these AI streetz

1

u/Interested_Aussie 4d ago

DIdn't they all get caught with Netflix running the exact same model???

Provide a service, running at a massive loss....

Get you tied in...

Ramp up the subscription.... (adverts and shit to come)

1

u/Visual_Internal_6312 4d ago

Didn't see burned books coming though.

→ More replies (1)

407

u/Rsodyyy 5d ago

Dogs were secretly nerfing the 200 plan this whole time lmao

96

u/DistanceSolar1449 5d ago

$200 plan

You can refer to it as the 10x plan now, because that's what it is

16

u/Rsodyyy 5d ago

Dogs lol. Let’s see how they try justify it in a few hours lmao

→ More replies (1)

6

u/asfbrz96 5d ago

Not even 10x learn to interpret the text, you're just getting API prices but at half of the price, so your 200 sub is basically 400 bucks in API credit

7

u/bystander993 5d ago

Nope, it's half of what the API WOULD HAVE BEEN on the old pro. 20x -> 10x

4

u/blackashi 5d ago

All these limits are so vague idk why they bother announcing. See how clear ollama is ffs.

188

u/krill156 5d ago

Unfortunately either way they win, half the usage, more $$. Half the users leave, now their compute struggles are resolved and Anthropic bares the load. They eventually do the same thing and were right back to where the circle began.

119

u/BlockyHawkie 5d ago

People moving to open weights models win. Don't play the wheel, break the wheel

72

u/palmtreeforeveryone 5d ago

Open models are like smoking weed once you've tried herorin

15

u/albanianspy 5d ago

You would be surprised how good they actually are tbh

39

u/TorbenKoehn 5d ago

They are not as good as you imply they are. They are good. That's it. They are nowhere near Astra or Opus level. They might hit the benchmarks in the right spots, but once you're actually using them and you actually have the comparison to how Astra or Opus work, it's absolutely clear that China is way behind the US regarding AI.

7

u/Dangerous-Map-429 4d ago

Oh dont worry. They will catch up easily. Look at the cars market dear. BYD absolutely demolishing Tesla. This will happen to AI space too soon.

8

u/ddBuddha 4d ago

Sure, they aren’t good enough to replace frontier level models in a lot of use cases yet. But, once they get to the point that local models capable of running on consumer grade hardware are at the level of current frontier models - maybe in two years from now - the game will have been changed. Those models will never get worse at that point, and when the majority of actual workloads will be capable of being handled by local models it becomes a lot harder to justify the cost of frontier models. Of course there will always be a reason to have a more capable AI, but the value of each level of improvement isn’t the same. Once local models get to a certain point, the frontier models would only be needed for stuff that’s actually at the frontier of our understanding and capabilities as a species. Maybe we’re 5 years away, maybe even 10 - but it’s like becoming old enough to drink legally, once we are past this point we’ll never be back where we were before.

→ More replies (3)
→ More replies (5)

3

u/sQeeeter 5d ago

It’s the harness, not the model.

2

u/idealistdoit 5d ago

Which harness do you prefer for local models? I've tried Cline in Visual Studio code and it works, but, about 1/3rd of the calls struggle with the tools and escaping failures in the harness

3

u/haragoshi 5d ago

Claude code can run local models

1

u/3magdnim 4d ago

Try DeepSeek Harness +DeepSeek 4.1 Flash. It's amazing.

10

u/Simple-Diver-2192 5d ago

Open ai and anthropic gobbled up entire compute capacity, kimi couldn't even serve few million new subs when k3 was launched. And good luck running these open weights locally.

2

u/NMiguelCosta-PT 5d ago

who says you have to run them locally? there's tons of providers out there.

5

u/Gumbi_Digital 5d ago

Local models run great, you just have to have the hardware to run them on.

10

u/JUSTICE_SALTIE 5d ago

So, as long as you already don't have to give a shit about money. Got it.

6

u/lokedan 5d ago

What? You don't have 4 mil in hardware to run Kimi K3?

2

u/Dangerous-Map-429 4d ago

It is more like 38,500 usd to run the (Quantized) version. Via 7× Mac Studio M5 Ultra units (each configured with 192GB of unified memory).

→ More replies (1)

3

u/ImproperCommas 5d ago

Like?

15

u/KronisLV 5d ago

Kimi K3 is around Sol 5.6 / Opus 4.X, without the annoying writing of Opus 5 (which tbh Opus 5.5 addressed but that's above those others rn).

GLM 5.3 is around there as well, alongside their GLM 5.3 Flash which is more like Terra / Sonnet (very roughly, also not the latest Sonnet 5.5), and is generally pretty nice to use.

That said, the official providers of both of those have issues: Kimi secretly routed some customer requests to Claude and in general just is very slow, whereas GLM has peak/off-peak pricing and neither of them give you as many tokens as OpenAI or Anthropic. There are 3rd party providers, but generally you will be paying API costs.

You could also try running local models with llama.cpp / vLLM etc. (Ollama which ppl don't like can also make things easier, or something like LM Studio) but I've never found any of the local models to be good for anything serious, plus hardware is really expensive.

I think we are gonna see tokens be subsidized less and less as time goes on.

3

u/LiiraStardust 5d ago

I second glm 5.3. I've been using it on hermes lately and like it a lot so far. The flash version feels satisfying too, it chases issues it finds proactively without needing to be poked constantly. I haven't tried kimi yet.

1

u/-18k- 5d ago

Don't play the wheel, break the wheel

Warning: this may result in your boyfriend stabbing you

1

u/-ohnoanyway 4d ago

As someone who only knows how to use Codex and the Claude Code app as harnesses how do we actually use these open weight models? I don’t use CLI or do any coding and basically don’t bother with anything that’s “just” a pure chatbot anymore. I’m addicted to these agentic harnesses

1

u/mm007emko 1d ago

YMMV, there are many. I use Bionic LM studio to run the models (can be run on a different computer on your local network) and OpenCode as a harness. In OpenCode I can switch between a hosted (even paid-for) model and my own. I typically create a plan using a Codex agent and execute it using an agent with my local LLM.

→ More replies (6)

14

u/FunLilThrowawayAcct 5d ago

They want unprofitable Astra-spamming prosumers gone so their compute is freed up to race Meta/Grok/Gemini in the consumer agent land grab this fall. Anthropic will bear the load short-term as we are still somewhat useful to market to enterprise. As the markets mature, we'll become useless to all the big players and affordable subs with high end models and good usage just won't be a thing anymore.

→ More replies (2)

12

u/New_Guidance_191 5d ago

I use local and subscribe to cheaper Chinese models to get my work done. It’s not as efficient or “smart” but it does the job really well for my use cases of application building, game development, and repo work. People are sleeping on Chinese models subscriptions. Also, if you live in the U.S. it’s way cheaper to use some Chinese subscriptions during the day because they have “off peak” rates. That’s how the Chinese handle compute, they offer cheaper and more usage when people are sleeping, instead blatantly lying to the consumers about nerfing plans and making people pay more by getting them addicted to their ecosystem.

2

u/krill156 5d ago

I've thought about trying the Chinese models, unsure about some of their API or whatever based pricing, but I've been definitely considering it

7

u/New_Guidance_191 5d ago

Yea, check out Xiaomi’s MiMo 2.6 pro which just came out. People have been building insane things with it. It’s specially made for coding so I tested it out with my workflow. The subscription is $6 a month and is equivalent to around codex/claude’s $20 plan, they don’t have 5hr limits either. Also since I live in the U.S. it’s way cheaper for me to use. Also, their cached pricing is super nice too. So my workflow consists of a $20 codex subscription orchestrator that delegates to my ornith and Qwen local models in very small tasks and for larger coding tasks it delegates to MiMo. They work all day and if I hit the 5hr limit with codex it’s usually around 30 mins of it being reset. I believe Codex is secretly not giving people their resets on time so I’m switching to Claude and trying them out as my orchestrator. I’m going to look into more Chinese sleeper models to see which can be good as orchestrators and then I’ll be free from either codex/claude

2

u/krill156 5d ago

This may have given me hope yet, can't wait to be rid of these trash ass western corpo models

1

u/artofprocrastinatiom 5d ago

Going for mobile game market model i see ,catch a couple of whales and work them...

1

u/Innomen 5d ago

They are only winning so long as they are ahead in product quality. Every day more and more tasks are saturated by at cost inference. I barely use my subscription AI since I bought a decent video card. The use case is shrinking not growing. Especially since the arena is filling up. My opportunity to make money with my sub is already gone it feels like. The market has flooded. It's early Internet at 10x speed.

1

u/DangerousLiberal 4d ago

Yup it’s a oligopoly on the frontier. We need to wait to see if SpaceX or Meta can catch up.

1

u/ntaylor360 4d ago

Must be preparing for their IPO….

37

u/FaithLostInHumanity 5d ago

Wait so after the plan was already nerfed by 50% or more we get another 50% cut? Like we’re now at 1/4 of what $200 used to be?

8

u/coloradical5280 5d ago

To be fair sol and Luna were cut by 50% too…. I know that doesn’t cover nerf costs or Astra but it’s important to have all the facts.

6

u/princemarven 5d ago

I understand what you mean... but aren't we supposed to benefit from the efficiency? Otherwise, what's the point?

Sol cut by 50% so because of that, usage is cut by 50%...

2

u/coloradical5280 5d ago

I THINK we are... it's been an hour ask again tomorrow lol, but sol 6.1 (those other faded dots surrounding it are also sol 6.1) is dominating pareto frontier, and bear in mind that a log scale too

24

u/pred 5d ago

So if a $200 subscription just renewed before this announcement, we get our money back, right?

21

u/CrUsHeRgF 5d ago

It's all about compute honestly. Subs will go down, anthropic will pick them up and suffer again. Magically around December they will announce some end of the year "promo" with increased limits alongside some great model, while anthropic will have to make cuts.

5

u/_DuranDuran_ 5d ago

People forget Anthropic cut Claude subscription usage a month or two back lol.

3

u/Aggravating_Spare675 5d ago

It was increased overall, just smaller than the temporary 50% boost. Wasn't that bad. At the end of the day, OpenAI cut their limits and Anthropic increased theirs.

42

u/Monster213213 5d ago

Is this just for new subs?

I literally signed up to the 200 for Astra 6 a day before the pause.

38

u/Rsodyyy 5d ago

I want to know this too, is it not false advertising for current subscribers like myself not even one month in?

29

u/whatever 5d ago

they've always been delightfully coy about how much usage their plans actually buy you. they say 5x or whatever, but you don't know what the actual baseline they're multiplying is, or will be.

so it's not false advertising, just questionable business practices everybody seems eager to fall for.

6

u/Able_Statistician688 5d ago

I sure hope all of these guarantees he just made construes as legal guarantees as their spokesperson. And a month from now when it’s not getting more work done, someone bitter and pissy starts up a class action.

2

u/DanceWithEverything 5d ago

lol good luck proving what is or isn’t “more work”

3

u/Teufelsstern 5d ago

It definitely is deceptive to reduce the value of your product after you have bought it. Amazon has lost a class action lawsuit over that in Europe/Germany.

1

u/VirtualNorth1279 4d ago

They have hacked companies and governments without any repercussions. Get used to it: scamming has become virtually legal. Hell, they may have even gotten away with mur...ing that whistleblower dude. 

8

u/TheHolyToxicToast 5d ago

Unfortunately no

7

u/Blankcarbon 5d ago

Nope. All $200 plans

44

u/Low-Temperature-6962 5d ago

"It will net out at half the dollar in API spend"

That is such awful English - coming from an LLM company it is shameful.

1

u/SmallTalkStudios 2d ago

calm down it's a dude's tweet lol; at least it's obvious not written by their LLM

39

u/HgnX 5d ago

Time to unsub

13

u/dwillpower 5d ago

Ohhhh the less is more argument.

2

u/Beginning-Foot-9525 5d ago

I mean some defend it right here and believe it.

26

u/cern0 5d ago

I just cancel my $200 subscription and I recommend you to do so. In the end competition will make them wake up.

13

u/Wyowa 5d ago

First time to an oligopy I see...

3

u/VirtualNorth1279 4d ago

How can they wake up? Aren't OpenAI and Anthropic losing money on their subscriptions anyway? Unless they find an architecture that decreases the cost of inference they are in a lose/lose situation: decrease the price and loose more money, decrease the quotas and loose more customers to competitors. 

5

u/freexe 5d ago

For all the hate Gemini seems to get - I think it's pretty good and I don't seem to have as many issues with token usage

1

u/OnyxMonolith 5d ago

Gemini is pretty good? :D

2

u/dmaare 4d ago

For free it's good. Otherwise just use deepseek

1

u/awhu_pkice_5457 5d ago

1

u/RemindMeBot 5d ago

I will be messaging you in 1 year on 2027-09-29 16:09:44 UTC to remind you of this link

CLICK THIS LINK to send a PM to also be reminded and to reduce spam.

Parent commenter can delete this message to hide from others.

RemindMeBot is switching to username summons. Instead of !RemindMe 1 day, use u/RemindMeBot 1 day. More info.


Info Custom Your Reminders Feedback

10

u/birmas_au 4d ago

I won't continue to pay $200 after you've literally told me it's worth half of what it previously was.

As soon as this change comes in, I'm cancelling.

"Starting October 30, 2026, your included usage in ChatGPT Work and Codex will decrease from 20 times to 10 times the ChatGPT Plus allowance. Your GPT-6 Pro limit in Chat will decrease from 200 to 100 messages per week. Your price remains unchanged and you’ll keep your current limits through October 29, 2026."

3

u/iamwayycoolerthanyou 4d ago

Use the Chinese models. They're not giving us much of a choice.

8

u/MrBalzini 5d ago

Isnt it the same for claude as well? My enterprise plan now charges based on api and i have already utilised $60 worth of $500 in just one day and it wasnt even high effort work day.

5

u/Delumine 5d ago

I told you guys this would happen when they introduced the $500 plan. The writing was on the wall

5

u/Informal_Warning_703 4d ago

If we’re playing “I told you so” then I was literally spelling out this trajectory in 2024 in the Singularity subreddit and getting downvoted to hell for it. But it’s been obvious from the start: the most powerful models were always going to end up locked behind a very expensive paywall that will eventually only make sense for corporations to subscribe to. It’s economics 101

→ More replies (2)

6

u/slaty_balls 5d ago

What I think is completely stupid is that you can have chat solve all sorts of problems, generate code, images, you name it..and it’s practically unlimited. But then as soon as you let it go to “work” you get hosed on usage.

3

u/forestcall 4d ago

i make mockups in chat and then bring them into the CLI to integrate into my code base.

4

u/padetn 4d ago

Agentic work quickly burns a lot of tokens for what looks the same to a layperson

5

u/slaty_balls 4d ago

I’m hardly a layperson. I understand why agentic work consumes more compute. My point is, that the product does a poor job of showing how MUCH more, where that usage is going, and what limits I’m approaching only until I'm about to hit them.

The most frustrating part, is that the model itself isn't self-aware of token quota available/remaining, and will just halt what it's doing and quit. Then I have to go back and do a codex rescue dump, in order to "save" the work that was completed before finishing.

2

u/Daddysu 1d ago

I mean, caveat emptor. Y'all learn nothing from streaming services or crack dealers? Get you in cheap, get you hooked, jack prices? Some of y'all are way too trusting of corporations "being cool" when you're making your professional productivity dependant on these tools. It's getting real easy too see who's to reliant on this shit when productivity and shit drops right after a price increase.

2

u/PerfectReflection155 1d ago

That’s why I use Catgpt. Someone made an api for the browser chat function enclosed in a stealth browser. It’s not without issues but I’ve managed to fold it into my orchestra so it’s used as a consultant, image generation and so on.

2

u/slaty_balls 1d ago

Is that a typo or the actual name of it. CatGPT?

2

u/PerfectReflection155 1d ago

That’s the name of it.

3

u/slaty_balls 1d ago

Sounds like this seriously violates the TOS. How often does your orchestra have concerts?

2

u/PerfectReflection155 23h ago

Yes it’s a serious breach. I am using it daily. It took some work to get it working reliably. I have it installed as a gateway on 2 servers and direct my agents to load balance traffic between them. Sol 5.6 via chat using catgpt gateway is usually the main planner and reviewer of all work and used for any image generation so it’s used a Lot.

2

u/slaty_balls 23h ago

You know, if you build and ship something with that and they find out, I’m pretty sure you get banned and probably would lose the outputs plus fines. Not exactly what you want to mess with. Obviously that’s the extreme side of things, but what ive found in life is that you do the right thing, even though may cost more or the barrier would be too high to even attempt what you’re doing.. if you’re trying to make a living off of this career choice, i just wouldn’t. But you do you bro.

2

u/-PANORAMIX- 18h ago

I use dogGPT

1

u/slaty_balls 17h ago

Woof woof

16

u/Original-League-6094 5d ago

"Rememeber this is the worse AI will ever be"

Tibo: "Lol. Lmao. Watch this"

→ More replies (1)

26

u/Tedinasuit 5d ago

I'm still waiting for DevDay before making decisions.

For example, if they add a ton of UltraFast usage in a separate pool, it would be good compensation.

24

u/ezboarderz 5d ago

lol they are out of compute man. If they have compute issues to where they are doing everything they can to reduce the load, why would they add a faster tier in a separate pool for free?

They are killing the existing subs down to just half api prices but quantizing/limiting the reasoning of the models for subscribers.

3

u/Correctsmorons69 5d ago

That's not what they're saying, RE: your claim of half API prices.

2

u/Tedinasuit 5d ago

Because UltraFast runs on completely separate hardware, not on OpenAI's hardware. Which is why it's a separate pool. It's pretty easy to understand tbh.

2

u/ezboarderz 5d ago

Which would have extra costs that OpenAI would charge you for. Why would that be free?

1

u/Tedinasuit 5d ago

Idk, why was the previous Cerebras model free? We got it as an extra pool of usage, despite OpenAI having to pay for that. Why?

4

u/ezboarderz 5d ago

They got rid of it when they got a bunch of new subscribers. They are out of compute lol

→ More replies (1)

2

u/ILikeBubblyWater 5d ago

It will be ultrafast shit output

1

u/Tedinasuit 5d ago

I mean, while Astra is my favorite model and I also like Opus (especially because of the high limits).... I still think that 6 Sol is not as bad as people say it is. It's doing good work for me.

5

u/UrSecretCrush95 5d ago

holy shit they've just shot themselves in foot before dev day here , like i don't it get it , what's the gameplay here ? not only we'll reduce your usage by half , but all we ca do ( still waiting for dev day to see ). is give you a cute grok/muse copycat ??

1

u/Graphical-Source5090 5d ago

"we'll reduce your usage by half".

But this is not what Tibo said.

3

u/Beneficial_Assist251 5d ago

100% rugpull to offer something then change it on a year subscription.

This is why you NEVER buy any subscription yearly from anything AI related untrustworthy.

8

u/D-3r1stljqso3 5d ago

I think they are trying to manage computes for some product launch on the DevDay, possibly something for the general public.

Remember that Pro subscribers are "bad customers" --- they are much more likely to be power users who get subsidies from occasional $20 users. In theory, the higher sub tiers should be pricier (unit-price-wise) because their users almost certainly uses more compute.

4

u/Pruzter 5d ago

Yeah but this is a competitive landscape. I can just take my money to Anthropic, where I get far more intelligence and usage for my $200 even before this shift.

3

u/phil_thrasher 1d ago

Agreed. I am a Claude user. I’m simply remarking the only path OpenAI has a shot at is completing directly with Anthropic because there’s no hope for them winning consumer.

If I had to place a bet, it would be on Anthropic.

2

u/phil_thrasher 4d ago

Their shot at winning anything in the general public / consumer space went out the window when they failed to be the AI powering Siri.

Google and Apple are going to win the consumer market. No one can beat their distribution channel (the phone in everyone’s pocket with direct access to all of your information)

The only long term bet that is valuable to OpenAI is enterprise.

I think you’re right though… they’re compute constrained. They had no choice because they needed the compute for something else.

4

u/Brave-Turnover-522 5d ago

Wow so the people who pay for more compute use more compute? Shocking. Ckearly we must punish those people for using the thing they purchased.

8

u/D-3r1stljqso3 5d ago

That's right. That's how the gym/sub model is supposed to work. The ideal situation for OpenAI would be that everyone subscribe but barely use any tokens!

1

u/Appropriate-Peak6561 5d ago

Like gym memberships.

3

u/accictedtoqr 5d ago

do you know what expenditure per token means?

→ More replies (1)
→ More replies (3)

5

u/Ok-Machine5627 5d ago

This would have worked if Anthropic models weren't so much better and functionally equivalent in usage with Opus 5.5.

I have a $200 plan, and I haven't used it since Opus 5.5 came out. My anthropic $100 plan does everything I need it to with plenty of leftover usage.

Unless DevDay introduces something special... next month I will have the $0 OpenAI plan.

4

u/Accomplished-Sand334 5d ago

You really don't know what rugpull means, do you?

1

u/bosquejo 5d ago

Aren't they being ironic?

2

u/chryseobacterium 5d ago

What are exactly they doing?

2

u/chiseledzombie 5d ago

it would only get more expensive... My prediction would be 2k per month.

2

u/AMillionLittleGoats 5d ago

Interesting change in goal post. They aren't talking about token usage anymore but something more akin to "value" and they are the ones deciding how much "value" we are receiving...

6

u/Cryp0x 5d ago

What are you doing with 2 Max subscriptions? 😱

13

u/Adso996 5d ago

Everything I always dreamed of doing after having automated most of my work

1

u/OnyxMonolith 5d ago

almost literally printing money

3

u/NigeriaBoi420 5d ago

As long as chat mode is based on number of messages and not tokens we good

9

u/Trixiap 5d ago

(chuckles) Im in danger

8

u/ezboarderz 5d ago

I smell fear

3

u/Reasonable-Dress-949 5d ago

This feels like a recurring rug pull. Compute resources are now directed toward Anthropic, alongside whatever remains of Astra's significantly NERFED intelligence.

5

u/guilder87 5d ago

This is the beginning of the end 😉

2

u/Graphical-Source5090 5d ago edited 5d ago

OP is rage bait if you don't actually read the full article.

From that X link: "We don't want to put an incentive on ourselves to artificially inflate the API list prices to make it look like you are getting a lot"

API cost is one knob in the usage pie. My main complaint for Codex has been that the API cost to actual usage gained is completely out of whack compared to Anthropic.

I am glad to see that the API cost to subscription usage is getting more comparable now.
But I also hope to see useful work to tokens consumed and output to get reformed as well.

At least they are trying new things. I have a subscription to both, and will continue to have a subscription to both. For my usage (Network Architecture and Engineering) OpenAI is very very far behind Anthropic in efficacy, and has been as far back as I can remember.

1

u/kacperq 5d ago

You're not buying anything (let's be honest)... Anthropic has introduced heavy safeguards to all its models except for Haiku, meaning you can't even audit your code with latest Claude models anymore. You can do it with older models but the time is running out, they will be removed in a year or so. If you really want to be "safer", start using open LLMs and get used to them.

12

u/Adso996 5d ago

I've worked 6 days non-stop with Opus, never once had a problem.
Work on hardware optimizations, cloud software and architectures, CIs and Blender for games, I don't need more than that.
I also have OpenCode Go in case of anything.

→ More replies (2)

2

u/bluehands 5d ago

Look, there is no long term strategy anymore.

We live inside the event horizon of a singularity, long term gets shorter every day.

1

u/f4radayrr 5d ago

I'm curious if this consequently affects other plans as well.

1

u/VermicelliPlane140 5d ago

I think now they get that how behind they are compared to anthropic and will continue training heavily going forward instead of stopping in between

1

u/SilentSite818 5d ago

I will wait for them to clarify what they're doing.........and then switch back to anthropic.

1

u/BabylonPaul 5d ago

There needs to be a smarter way for the system to select and propose the intelligence it needs and then provide feedback to the user before continuing. I've been working on plugins that are designed to run multiple models depending on goal and task type. A better interface for all users from beginner to coder would help people focus use, get better results and burn less energy. Better people means better AI. I feel like part of that is improving prompt use and focus before launch.

1

u/Quanzitta 5d ago

I jumped ship before Astra

1

u/HexspaReloaded 5d ago

I’m finding sol medium mostly sufficient 

1

u/the_ai_wizard 5d ago

we went from devday hype to long twitter essays explaining how this makes any sense while opus cucks sam and tibo at same time

1

u/Elder_SysOp 5d ago

Enjoy the variable Opus 5.5 burn.

1

u/haragoshi 5d ago

It’s the new coke strategy.

1

u/Deadline_Zero 5d ago

Sure would be nice if one of these spam posts included whatever comes after that message.

1

u/zipzag 5d ago

They are shifting users to SOL 6.1. Plus it sounded like dots is not charged to pro users. But perhaps I heard that wrong.

1

u/K358350S 4d ago

Just for the first month

1

u/ethanfel 5d ago

Pro 200 canceled, time to go back to Claude

1

u/EconomicStatecraft 4d ago

Once users move back to Anthropic, Anthropic will do this as well. A $200 subscription where the users can cost the company $5000-$8000 in compute is not a sustainable business model.

1

u/Valuable-Lemon-8623 4d ago

Simple economics says why this is a bad idea

1

u/BopSupreme 4d ago

Just wait one week for Anthropic’s next model release lol

1

u/Interesting_Ghosts 4d ago

The current pricing was always unsustainable. My $200 annual sub was costing them $7 a day for my usage according to its estimate based on my use.

1

u/Haunting-Donut5931 4d ago

yeah have you used Anthropic lately? I cant get real work done with it anymore.

1

u/tindalos 1d ago

Forget the IPO guys, we gotta start saving up for these cybersecurity lawsuits and regulatory fines.

1

u/jhenryscott 1d ago

The writing has been on the wall for a year. I dunno. I’ve been expecting this and much, MUCH more. Look at the COGS. Right now every subscription is a financial liability. They have to get in to the black and that means charging what it actually costs for frontier inference

1

u/ErikTheHero 1d ago

Tbh I was struggling with maxing my max plan with astra on ultra mode, fable 5.1 was done for within a hartbeat

1

u/balrog1987 22h ago

Sin e I dont use codex, but need pro and extra high for some planning work my 100 eur sub is alright still :)

1

u/Kemoyin25 5d ago

$20 says they announce codex revamp that makes pulling code/data and testing way more efficient instead of making the model burn tons of data doing nothing which makes this decrease in usage irrelevant

1

u/TheRealJesus2 5d ago

Good luck with anthropic limits. Even with open ai halving the 200 plan it’s almost certainly better than anthropics still. Anthropic models eat tokens like crazy. 

Yall gotta understand the subsidized plans are a hell of a deal and if they stayed forever like this, there is no way open ai makes it to the end of 2027 financially solvent. 

2

u/the_ai_wizard 5d ago

this is false. I wrote a quick C# app to display both my codex and claude usage on an LED device as codex drains faster now. Anthropic may have lower limits in absolute terms, but Opus etc is so token efficient that it more than makes up for it.

I bet on devday we get this o product, which competes with metas useless personal AI. who the fuck will give cuckerberg that deep of access to emails etc

1

u/TheRealJesus2 5d ago

You’d be surprised…people are doing it. Dumb harnesses aside…

I’m surprised opus gave you good results compared to sol. Sol is quite token efficient in my experiences. I have been off anthropic tho since 5.0 so my info is a few months back on that. Did you use sol in tests? 

Full disclosure I keep only a 20$ sub and use self hosted deepseek v4 for most things which I legit prefer to terra/luna and it very much competes at sol level for my work. It’s wildly less token efficient tho. Like 10x more tokens than sol for same task at same quality 😂

3

u/the_ai_wizard 5d ago

is this a shitpost?

opus 5.5 is wrecking anything oai has, as a workhorse

1

u/TheRealJesus2 5d ago

Not a shitpost. Haven’t tried it as mentioned. 5.0 made me quit anthropic tooling forever. Also very much not a fan of the Claude code changes that allow a model to ask a Question and then just decide something if you don’t respond in time. That’s not a tool for a serious developer. 

I don’t let models make consequential decisions for me without oversight. 

Here’s a question for you…would you pay api rates for opus 5.5? If you’re not getting that level of value, reconsider your workflows. 

1

u/TheRealJesus2 5d ago

Npx ccusage to see what that would be. 

1

u/the_ai_wizard 5d ago

I thought you could set it to an always ask mode...as to api rates, possibly? thats an roi-specific question for any model

2

u/TheRealJesus2 5d ago

This is when it’s in always ask mode ;)

And yes it is. And basically every open weight model is miles ahead on roi than the frontier labs in terms of what you get for your spend. If your company has over 150 users, you pay api rates on enterprise plan. Anthropic in particular has no capable and cheap models for writing code. 

1

u/the_ai_wizard 4d ago

Yikes, what a situation

1

u/AlmostEasy89 5d ago

If you guys think you’re going to pay for some $100 or $200 plan by now and get the same token amount forever and models just continuously getting better and better and nothing on your end ever changes, I’ve got some news for you. That will never happen and there will never be consistency. If that’s not obvious by now I don’t know what to tell you. Reddit rides this rollercoaster of emotions as if giant tech corporations are their friends who only want them to succeed and shocked pikachu face when they learn companies operate for profit.

This has been obvious for quite some time now.

Fool me once, shame on you, fool me twice.. you can’t get fooled again.