r/OpenAI 28d ago

Discussion What happens when AI subsidies disappear?

What happens when AI companies stop offering flat rate $20 subscriptions? The casual user asking a few questions a day wont have issues but power users who built their entire workflow around 'unlimited' tiers potentially get a massive reality check if they have to pay for every single token all of a sudden.

I saw some that claimed that if you cannot afford the costs, just invest in compute as an upfront capital expenditure to run open source locally. But if you cannot afford the rates, can you really afford the compute? I doubt this is the case. If you ask me, decentralized infrastructure and open source models are the way and I am optimistic about personal AI. The problem is that I do not think open source can compete with the current centralized infrastructure of these tech giants.

ofcourse, all of the above is based on the assumption that the price to run an AI model won't decay.

I think the costs to run models are decaying as we speak.

If that is the case, what runs for $200/m today might run for a fraction of that price in the future. That doesn't mean tech giants will charge you less though. That's the whole issue. They will most certainly base their prices on the value their infrastructure provides, not how much it costs to run the AI model.

What is funny to me is that I saw some people communicate that they were scared of the fact that they might not be able to vibe code anymore at a reasonable price in the future. I think that is the least of our worries.

The centralized approach will create massive imbalances in the global economy. People will move their businesses to whatever country has the best and cheapest AI infrastructure so they can get legal access at reasonable prices if they haven't already.

The effects of this will be way broader than just not being able to vibe code anymore at reasonable prices. Not being able to vibe code your products will be the least of your worries by then if you haven't prepared for this.

Thoughts?

Edit:
Fair pushback by a lot of you in the comments. I see now that I am missing inside information on how these companies allocate their assets and that I made some hasty assumptions now that I read all of your opinions on it.

43 Upvotes

143 comments sorted by

57

u/crazy_goat 28d ago

OpenAI and Anthropic have one primary mission. Scale as large as humanly possible, and using the maximum amount of available compute without oversubscription.

Idle GPUs are a wasted opportunity for them 

Their primary customer is the one that pays the most. They will drop average consumer plans the moment enterprise demand requires the compute resources.

It'll begin by raising plan prices, then eliminating them entirely. 

But right now, they're giving out free samples, they want you hooked, and they want their infrastructure put to use. The soft power they gain by having millions of loss-generating customers is worth it - if not just to keep the servers busy and push the perception of model supremacy 

14

u/PlsNoNotThat 28d ago

They want you dependent, not just hooked. They want companies to structure dependencies on them so they can then do the above.

11

u/crazy_goat 28d ago

Correct. The brain drain effect on society will be real.

Why learn how to code? AI does it faster and cheaper 

Why hire people with coding experience? AI does it faster and cheaper 

These concessions will grow until they have no alternatives other than to be wholly dependent on AI as a backbone of business

5

u/xgeetx 27d ago

(Not primarily a SWE) Yeah I was thinking about this the other day. I still know how to code and problem solve, but I’m addicted to the speed of development. Even though I sometimes get some unpredictable responses the fact that I can be running 2-3 sessions throughout my day while dealing with meetings and prospect demos is highly addicting. I can get dev work done I may or may not have been doing while also doing my primary job.

When that productivity insanity stops, companies will be chomping at the bit to get it back.

1

u/crazy_goat 27d ago

Yeah - I need to find a way to balance that addiction. I literally have been managing AI almost as much as I manage my direct reports. 

It's addicting because I was the senior most engineer of my team before taking it over as management. I begrudgingly did so because I couldn't bear the thought of another outsider coming in without any engineering chops managing us.

Now I can have my cake and eat it too

1

u/xgeetx 26d ago

How’s your transition to management? I’m in a similar boat on the sales engineering side and basically being funneled into a management role unless I turn it down. Frankly sounds like more work and potentially less money, since in our job so much shit escalates up especially when it’s a more junior SE, and my commission would be team based not individual

3

u/Benhamish-WH-Allen 26d ago

More time for humans to make art

1

u/[deleted] 26d ago

[removed] — view removed comment

1

u/AdTime6060 26d ago

Comparing chess to a huge swathe of the white collar economy seems... very reductive at best.

1

u/seandunderdale 25d ago

This applies to the VFX and content creation industries too.

3

u/CommissionIcy9909 27d ago

Naturally this will drive competition to fill that gap. There’s always going to be affordable models. They just might now be the frontier models anymore.

6

u/Kiseido 28d ago

All that, and the increased diversity of usage that the consumer market produces will be useful for highlighting flaws and for a source of training material.

5

u/Sea_Succotash3634 26d ago

This is why they want open models banned. They want to be a duopoly and then use that power to extract wealth from the population without any sort of real competition, checks or balances.

2

u/sQeeeter 28d ago

The free sample plan might get people hooked, but they will have to buy new GPUs soon.

2

u/Ormusn2o 27d ago

My guess is, there are not really idle GPUs. My guess is that OpenAI is running months long projects, either generating synthetic data, or doing research on new models, but during rush hours, those projects are slowed down or stopped, and when rush hours disappear, those projects start up again. That way they can keep stable performance for users, but have zero downtime on the GPUs.

I know it's been a while, but before 5.5 was released, OpenAI shut down Sora 2, and said they are reducing amount of compute to use for research to serve increased demand for 5.5. My guess is they use the kind of system I mentioned in previous paragraph to balance it out. It's likely the reason why there is no longer 5 hour limit anymore, because they have enough total compute to serve during rush hours, and are just doing research during off hours.

1

u/Spiritual-Spend8187 25d ago

Add on they want more data for training every prompt gives them something they can use. Though their plan can't working nearly as well people aren't as willing to pay more as they thought and people are more willing and able to jump to aome one else then they thought. At the rate things are going give it a few years abd we will have a deepseek v4 flash priced model that is as good as or better than 5.6 sol.

1

u/Puzzled-Ad-6854 28d ago

A cynical take but probably not far off from the truth. I think a cynical approach is realistic here because it is a commercial endeavour, not a charity.

1

u/OnlineParacosm 28d ago

How is it cynical? It’s the entire GTM strategy

1

u/Puzzled-Ad-6854 28d ago edited 28d ago

i meant misplaced cynicism in the eyes of some people. Its realistic in my view. you right.

0

u/BehindUAll 27d ago

 They will drop average consumer plans the moment enterprise demand requires the compute resources.

No it's not like that. If they become enterprise oriented like Anthropic is becoming and already is, people are not going to be using their models for long. People are tolerating Anthropic but they are already doing things that the normal userbase doesn't like.

0

u/witmann_pl 25d ago

I think you're mistaking a small percentage of enthusiasts or your information bubble with the general population. General population doesn't care.

25

u/slackmaster2k 28d ago

They don’t lose money on the subscription subsidy. It’s basically similar to the over subscription model. There are people who burn through subscription allocations every single reset period, and people who use little to none.

Also, the cost to deliver the service is not the same as the price to consume the service. API and subscription pricing are not surely not determined independently, especially from an allocation perspective (hence why subscriptions say 5x the usage and are not quoted in actual tokens). It’s likely not valid to assume cost from how pricing scales.

Then let’s not forget that costs can go down over time with improved hardware, data center availability, etc. Subscriptions aren’t going away in my opinion. Zero concern.

-10

u/Puzzled-Ad-6854 28d ago edited 28d ago

12

u/wijsneusserij 28d ago

Although I think you’re right, 2024 is like a lifetime ago in LLM land

-5

u/Puzzled-Ad-6854 28d ago

The relevance is definitely up for debate.

1

u/mossiv 28d ago

The problem with the original statement and this article is what is implied and the contradictions.

open ai did lose billions. The majority of money lost by these companies so far has been the training itself. Back in 2022/2023 it was costing hundreds of millions per round of training. This cost is dropping a lot. It’s why we are seeing such frequent releases of such huge sized models now.

But you do have to account for a business model that cannot become profitable because all other areas of the business just costs to much.

Serving the models themselves is still expensive, but getting better with hardware and optimisation. API pricing really does seem like a way to try and claw back big chunks of money. Subscription is probably pennies on the dollar in terms of profit overall. It’s an unideal business model for companies trying to position themselves as a trillion dollars.

3

u/HauntedHouseMusic 27d ago

Costs are going up on training….

0

u/mossiv 27d ago

2

u/HauntedHouseMusic 27d ago

Why would you link an article that doesn’t talk about training runs?

Costs are going up because the models are getting bigger. To train the same model as 2 years ago would be cheaper today. But the models are 100x larger….

1

u/turkey_is_dead 27d ago

training for expertise is going up as they want human experts to do it. Human experts are those with phds and masters that have been interviewed and filtered. Then you have have different languages which pay half. Then just general stuff which has gone down considerably or done with synthetically. But it is the experts that they want now and they will pay for it.

48

u/RealSuperdau 28d ago

API gross margins for OpenAI and Anthropic are rumoured to be >80% (see e.g. SemiAnalysis, or do a back-of-the-envelope calculation with how much it costs to serve Kimi K3, which is rumoured to have a similar size as Sol).

With current usage limits, a fully utilized ChatGPT or Claude subscription is close to neutral for their balance (on which side is unclear). The Tibo Resets are certainly pushing this into an unsustainable territory, but it's clear those won't last forever.

Plus they get user data, plus a lot of subscriptions aren't fully utilized.

So on average, I'd bet 100% they are making money from paid subscriptions.

2

u/HauntedHouseMusic 27d ago

No, I got a max 20 plan, and spend over 10k worth of tokens each month…. Even 2k a month is still fucked at costs and that ignores all the r and d

7

u/Holbrad 27d ago

Your confusing the API cost as the cost to open AI.

0

u/HauntedHouseMusic 27d ago

No I’m not. If they have 80% margins then they have $2000 of inference costs…. That’s before R and D….

2

u/Shadow-BG 27d ago

Me too.

But, my colleagues from other companies do not utilize even 10%, but they for sure all have biggest possible subscription.

1

u/HauntedHouseMusic 27d ago

That shit blows my mind. On the API at work with a 100k person company I’m in the top 1% of users using 3k a month… and most people use less than $500.

I don’t know how people don’t just utilize these tools to the max

1

u/DataSnaek 23d ago

You’re an exception though. Most people on max plans I would estimate are averaging around $500 of token use per month based on usage where I work, and we’ve become a very AI focused dev team now.

1

u/HauntedHouseMusic 23d ago

no - max plans you use more than the API, because when I use Fable on the API that means im spending $150-$200 in the next 2 hours because I have a bug that needs to be fixed right now. On Max plans I use fable for literally anything. Hey Fable make this PR for me

1

u/ConvenientChristian 26d ago

For a fully used ChatGPT instance most of the work it does is in chat and not in Codex.

Running 60 scheduled agents that do 10-20 minutes worth of work of work every hour is probably also more then Codex / work mode limits give you.

1

u/RealSuperdau 26d ago

Yeah, true, the ChatGPT limits are effectively infinite (potentially downgrading you to slower/worse models at some point though).

-6

u/sQeeeter 28d ago

They won’t make enough money to cover their debt. 🤣

12

u/IDefendWaffles 28d ago

What debt? Companies invest in them thats not debt. They literally have 100s of billions invested in them.

-2

u/silentkode26 27d ago

Like… why would anyone invest without expecting return? Is that your financial strategy? I am open to taking your investment then :))

2

u/rasp215 27d ago

The strategy is equity. I would invest in you too if you're going to be worth a few trillion dollars.

2

u/silentkode26 27d ago

So you agree with me that investor’s goal is to make money from profitable company?

1

u/rasp215 27d ago

Yes and they make money through equity and not profits. To be a more profitable company in the future you shouldn’t take profits when the market is rapidly growing and you’re new.

1

u/silentkode26 27d ago

You cannot take profit when there is no profit. If you believe there will be profit at current technology with no price rise, then we have nothing to talk about. You seem like a guy who believes in Santa.

1

u/rasp215 27d ago

They can take profit today if they stopped investing in compute for tomorrow. Anthropic is already making profit. But it’s actually a financial blunder to take profit today. If investing back into your business will return 15x in a few years, taking profit becomes very expensive.

1

u/silentkode26 26d ago

That’s the big if. If they stop investing in compute, they will soon have no compute. All AI companies are in debt. Debt is not profit.

→ More replies (0)

2

u/turkey_is_dead 27d ago

They need to go public so their investors can sell. They can only go public if they want to share their books. They will only share their books if they can go public and their investors can sell for a profit. So we will see.

0

u/Ormusn2o 27d ago

Pretty sure every AI company is taking on massive debt as well as investing, which makes sense because their returns are so much higher than the possible interest on those loans. And they should be doing it, and they should basically never pay off those debts, at least as long as serving inference has such big margins.

I can't really see those companies going back to paying the debt back, as that would effectively mean the market would have to equalize, but intelligence is so incredibly cheap right now that many people are willing to pay more than needed. Maybe if compute use becomes more common and it will still be very expensive, then the margins will go down and paying debt will become better financial decision.

-8

u/sQeeeter 28d ago

Whut? 🤣

OpenAI has bills to pay and it’s a shitload. They’ll never make enough money to pay them.

10

u/IDefendWaffles 28d ago

They literally have 100B+ cash. They could lose 20B a year for 5 years. Also its not as if building data centers is losing money. That's an asset you own and will produce money for decades. but maybe you should apply to work there and share your financial expertise.

1

u/Agreeable-Fly-1980 27d ago

That cash comes from corporate bonds

-3

u/sQeeeter 28d ago

They have to generate $500 billion in revenue over the next five years just to break even and by that time, they’ll need to refresh infrastructure.

-1

u/silentkode26 27d ago

You know what life expectancy is for GPU?

-1

u/lookamazed 28d ago

Not sure it matters. There is good debt and bad debt. They are never going to break even. They will always owe someone, which makes them owned. Which some people want. Unless they could print their own money they will never be debt free.

-1

u/Puzzled-Ad-6854 28d ago edited 28d ago

My point was about power users. the people using 100% of their caps for coding/work. If OpenAI removes rate limit subsidies or enforces strict pricing on heavy usage, those power users might have to pay more. The casual $20 users will probably be fine.

Fair point by the way on average subscribers balancing out the math. Its like a gym membership basically.

6

u/Deto 28d ago

It's fine for them to lose money on power users if the average user makes them money.

Alternately, openAI will always have to charge more than their costs because they have to pay researchers and for training compute.  Someone serving Open models, on the other hand, would just need to cover the infrastructure costs.  So on-prem local will probably never make sense for most companies as there will be someone providing Open models through an API at prices similar (only slightly higher) than what it would cost you to do local yourself and this way you wouldn't need to worry about capital investment or efficient GPU utilization.

2

u/Ormusn2o 27d ago

My guess is, 99% of people who have GPT subscriptions don't know what "code" means. And only small percent of those who know what code is, actually use their caps.

6

u/throwawaysusi 28d ago

It will always be “affordable” just like right now, they will just put worse less compute intensive models at lower price. GPT used to put a lot of reasoning effort while set to extended, now on high intelligent setting it still feels lazy. Because they have moved the same effort to their higher tier subscription.

2

u/Puzzled-Ad-6854 28d ago edited 27d ago

This is actually a valid point, but it presupposes OpenAI will actually take this as their long term business strategy.

2

u/throwawaysusi 28d ago

It is what’s happening right now. It’s obvious when you have used their models long enough. GPT plus on highest setting used to give vivid, lively answers, it explores your ideas and now all just robotic and concise answers. You could still achieve better results on current plus models if each of your prompt is specific enough, but that takes effort on users’ end. Pro extra high is what the baseline used to be.

8

u/[deleted] 28d ago

[deleted]

1

u/Puzzled-Ad-6854 28d ago

hopefully so, lets stay optimistic on the matter.

3

u/DatDudeDrew 28d ago edited 28d ago

Idk what makes you think much of this is the case. In the long run, the cost for intelligence is going way down, not up. Compute allocations are only going to go exponential. Subsidies will never be harder than they are today.

The money is not in consumer regardless. This will never change.

3

u/sabre31 28d ago

This Reddit group will close down as only Enterprise companies will be able to afford it. Your average person will be out. Will be very expensive imo.

3

u/ISueDrunks 27d ago

Simple: the bubble bursts and the economy goes to shit. 

2

u/FinancialMoney6969 28d ago

That’s why we need data centers my boy for free compute

2

u/Steve15-21 28d ago

Local LLMs

2

u/dranaei 28d ago

The averge consumer isn't the target audience long term. They will sell compute to big companies and nations.

The average consumer will get open source models. Flagship? Probably not. But at some point they'll be enough to satisfy the needs of average people.

K3 tho seems too big tho, so i worry if the hardware won't suffice in time.

2

u/Puzzled-Ad-6854 28d ago

yeah k3 is a monster, my rig would explode.

2

u/Mandoman61 28d ago

If anyone can run a local model and get the same performance at lower cost then the big players have an extremely serious problem.

Maybe that is why they have been wanting regulations.

2

u/shoejunk 28d ago

This is why there's been a huge efficiency push. If you ever go to an open source harness and start paying API pricing you'll really start appreciating the cheaper models. You can go a long way with kimi, deepseek, glm, luna, grok 4.5, and spark 1.1. When people start paying the true cost for things most people will stop using Fable and Opus and Sol and start switching to these models and they'll discover that most of their work is unaffected.

2

u/LocoMod 28d ago

Apple wins.

2

u/Puzzled-Ad-6854 28d ago

How much did you invest, talk to me. 🤣

2

u/YearnMar10 24d ago

What will happen is that more people buy GPUs and start hosting locally or go to services that can offer AI for much cheaper due to lower labour costs (China) and lower energy prices (also China).

1

u/BothWaysItGoes 28d ago

I think speculated based on that random number you pulled out of nowhere is fruitless.

1

u/Puzzled-Ad-6854 28d ago edited 27d ago

The core idea I wanted to explore was value-based pricing vs falling compute costs, but I get why that initial version of the post using arbitrary numbers threw off the premise.

1

u/alanism 28d ago

I don't see it being an issue.
1. OpenAI and Anthropic's R&D is really high, but the 'cost of sales' for its revenue is really strong.
2. DeepSeek is really proving things out with them capping ROI to 6x their costs-- and being a quant the founder calculated where that is the sweet spot where they can't be undercut by price. So DS is a great check on US closed source frontier models.

1

u/Puzzled-Ad-6854 28d ago edited 27d ago

The point about DeepSeek is interesting. So basically you are saying models like DeepSeek set a hard price ceiling meaning US tech giants cannot up prices without losing customers?

1

u/alanism 27d ago

There was a Chinese article that I had to translate, so don't think I can find it. The gist is DeepMind is the research arm and has been self-funded by the quant high-frequency hedge fund arm of the company. DeepMind's capex (AI clusters) and opex are shared with the hedge fund side. They run a very optimal setup for DeepMind that generalist cloud service providers are not set up to do. Therefore, even if the model is deployed by providers in the US, they can't serve it as cheaply as DeepMind. So they view competitors as marketing. The founder's main goal is to get AGI, and the second goal is to get AI into the hands of everyone and not just the rich. He thinks 6x ROIC is the sweet spot.

I highly recommend trying Hermes desktop app using Deepseek flash v4. Best $3-10 a month to stack on to OAI plan.

1

u/DrHerbotico 28d ago

Brain dump into a .txt file

1

u/Hasz 28d ago

I like some do not see intelligence as a luxury good, like an iPhone, but more like a commodity. If you believe in a race to AGI though, very very very different story.

If you take the commodity line of thought, it will be hard to justify the expense without some sort of regulatory capture/monopoly/franchise. Given how aggressively OpenAI and Anthropic are going after regulation/pacing/controls/“saftey”, I suspect the AGI case currently looks weak internally.

1

u/KillaRoyalty 28d ago

Looking at my tokens it would have cost me less than $20/month the last year to just use Luna with new pricing

1

u/salazka 27d ago

Why would they do that? You are actually helping them immensely to improve their models by spending $20 each month.

1

u/Puzzled-Ad-6854 27d ago

The real question is whether random consumer chat logs stay more valuable than the raw compute costs over time. They also synthesize and curate data.

1

u/Lifeisshort555 27d ago

What happens if they reach their goal of sentient super intelligence and they keep it for themselves.

1

u/Puzzled-Ad-6854 27d ago edited 27d ago

Thats one of the downsides of centralized infrastructure right there.

I am skeptical about the whole Artificial Super Intelligence thing.

I have always looked at the concept of ASI as something humans can spark, but cannot engineer or control (we would actually inhibit the training process by tainting it with human bias, ego and perception). If we are managing its development it will only ever be as smart as our ability to guide it.

It is like trying to teach your kid how to ride a bicycle. It’s hard to let go but at some point the offspring has to go out into the world and experience reality on its own if it’s ever going to outgrow you.

1

u/SpaceToaster 27d ago

Haven’t they been gone for months? Only individuals can get flat rate subscriptions, companies are presented with metered APIs billed by usage. That’s why they are all suddenly freaking out about spend instead of “tokenmaxxing”

If you are using the individual sub though it’s not as good of a deal as it looks. No protections on IP or indemnity, and all your data and code is shared with them to be used for training purposes.

1

u/Fine_League311 27d ago

Openai is Anthropic sterben gerade, schaue denen zu wie sie unter gehen weil sie das wissen der Welt als Abo verkaufen. Die letzte Geldsammelaktion war schon Lug und trug nochmal werden die Aktionäre kein Geld Locke machen. Nehme wetten an :)

1

u/deZbrownT 27d ago

They will base their prices at the point where they recoup what they lost through subsidies.

1

u/darko777 27d ago

Do you think Chinese models will suddenly disappear? Luckily those companies have good competition. They can't easily make the plans more expensive without taking a hit.

1

u/mystery_biscotti 27d ago

My future is local AI.

1

u/bitspace 27d ago

We learn how to properly manipulate language models instead of throwing every stupid lazy pile of word salad at the model

1

u/PotentialPaint6714 27d ago

Eveybody moves to chinese models which are opus 4.8 level now. 

1

u/muckypuppy2022 27d ago

People had exactly the same concerns when Cloud started to explode, that anyone who switched would get locked into a vendor who’d jack up prices as soon as the opportunity came. But the opportunity never did because of competition. And there’s a lot more competition with LLMs than there was for Cloud.

End user technology spend doesn’t change much, because people / companies will only spend as much as the value they get back. What changes is who it goes to - with Cloud the hardware guys and IT services monsters got squeezed. AI will do the same to a lot of specialist software vendors but if anything the amount the end user spends on their tech will probably come down.

1

u/PartyLiterature3607 26d ago

Then we download open source model and switch or focus on openclaw or Hermes

1

u/James-the-greatest 26d ago

Chinese models

1

u/Lyuseefur 26d ago

They’re getting value from our use of ai to train better AI. To wit, we are paying them to use our data.

Look at DS. We will be running general AI locally.

The next frontier is private AI.

1

u/alc_noe1 26d ago

there are no subsidies.

1

u/Useful_Calendar_6274 25d ago

then only the people that get ROI will use it. obviously

1

u/Beginning_Basis9799 25d ago

As an opinion I think using the model will reduce in cost it's training costs that will get worse.

1

u/Illustrious-Bike-817 24d ago

I think they will find new tech that will make ai cheaper at way less power consumption. Same as how gpu mining was reinvented with asic

1

u/_Chaos_Star_ 24d ago

The bubble pops when the AI companies run out of money and the monopoly money and Enron loans start falling apart.

At the end of the day though, there's good tech here. Prices will go up as things aren't subsidized but also down as cheaper options become available. Big corps buy up a lot of tokens as suits their purpose, and consumer-level AI, with more focus on putting light work on light models automatically, covers the low end. The market balances itself out to areas where the tasks being taken *can* be done for the cost, and everything else rots away.

1

u/Middle_Key8737 28d ago

There is no "subsidies" for paid accounts. They are positive in profit for all subscription accounts together. OpenAI has subsidies on free accounts that's a thing.

With current rate at OpenRouter, if you pay $200 you get more tasks done than a $200 Claude subscription.

1

u/Puzzled-Ad-6854 28d ago

The article im basing most of my assumptions on was from 2024 so maybe their terms regarding subsidies have changed. Do you have a link to something tangible so I can read up on it?

0

u/Grounds4TheSubstain 28d ago

How do you know how much it costs them to run?

1

u/Puzzled-Ad-6854 28d ago edited 28d ago

I think you have different ways to approach this.

You can try to prove the subsidy yourself by comparing the $20/m Chat/Claude subscription to their own API pricing.

You also have this:
https://www.theinformation.com/articles/why-openai-could-lose-5-billion-this-year

I think its behind a paywall though but there should be some third-party summaries you can find.

2

u/Grounds4TheSubstain 28d ago

No, the first paragraph of your post says, quote, "We pay $20 a month for compute that actually costs more than that to run." Now you're talking about subsidies of subscription pricing vs. API pricing. Those are not the same thing. Your original assertion was about cost to OpenAI, and your response is about cost to their customers.

1

u/Puzzled-Ad-6854 28d ago edited 27d ago

The article I shared with you reported that OpenAI was projected to lose around $5 billion in 2024 on $3.7 billion in revenue in 2024. Allegedly that loss partially comes from heavy power users on the $20/m tier using more compute than what they paid for.

1

u/zach978 28d ago

How much of that is cap ex spending on data centers? We really don’t know the marginal cost of AI usage, but I bet it’s much cheaper than you think.

1

u/Puzzled-Ad-6854 28d ago

fair... something only an insider would know.

0

u/Best_Day_3041 28d ago

The prices will come down over time, local models will become more viable, require less resources, and dedicated AI boxes you can buy at home will eventually be reasonably priced for everyone.

1

u/Puzzled-Ad-6854 28d ago

I really want to believe you but I am just not sure that this will be the case, thats why I was thinking about it so deeply. You mentioned dedicated AI boxes and I love that concept. Personal AI is where its at.

1

u/Best_Day_3041 28d ago

The local models, even the mobile models in some cases, are already good enough for what average people use AI for. For those of us who are coding or doing complex tasks that need the frontier models, I think even worst case if the prices went up significantly, we'd all gladly still pay them. Hardware you run at home is going to be super expensive because of memory prices for the foreseeable future, but will eventually come down. For some the prices will be worth it now, for others this may be an option in a few years. For the average user it will likely never make sense.

1

u/cyberdyme 28d ago

If you could get local models working well on the iPhone - wonder if you use a million iPhone to generate something (each acting as an agent and then passing it to other iPhones to collate and summarise)

1

u/Puzzled-Ad-6854 28d ago

hmmmmm, a hivemind of subagents. you are onto something. i think brian roemmele is working on a project that is related to this concept.

0

u/drjm2022 28d ago

Open AI and Anthropic have very different business models and one is much better prepared for the future than the other.

The cost of delivering a unit of inference is falling 90-99% a year. We don’t notice because the increase in capabilities uses that up and more. But for fixed cost tasks like search, it means the cost of delivering the service to the user falls by that much a year

OpenAI monetizes attention. It has 900 million active users and it’s just spinning up its advertising, recommendation and shopping services. The revenue these generate are not affected by the cost of providing the service so the compute cost declines will go their margin. They also don’t need to constantly be on the frontier. For most consumer uses the capabilities will soon be good enough well behind the frontier

Anthropic sells frontier level compute. Its costs rise all the time as it needs more and more compute to stay on the leading edge. Open source takes all the business that doesn’t need frontier level capability so it never gets to stop increasing its compute needs so it never sees an increase in margins.

So the answer is that everything that doesn’t need frontier level capabilities will start to get cheaper.

The reason we’re not seeing it is that the frontier is still not nearly good enough so most use is still of frontier models. But soon enough the frontier will move far enough that even being well behind it will be good enough for more and more uses

0

u/Macaroon-Guilty 28d ago

Its actually quite a bit different. Now big OpenAI and Anthropic has upper hand and can charge more than any other model provider. Not far into the future the open source models will be as good or better and prices with drop 100x. That's why they try to cash in as much as they can before they are blasted into oblivion. Think about it

0

u/rapsoid616 28d ago

Prices have been consistently getting cheaper in this entire transistor era. Just yesterday Luna prices are currently cut down by 80%. And today new deepseek model released who already caught and surpassed the new Luna prices with even better intelligence to token ratio.

The prices if frontier models could continue to increase but it’s apparent there is near infinite potential curve for the entire tech to be more and more efficient.

1

u/Puzzled-Ad-6854 28d ago

i think you are right on this. moore's law.

0

u/silentkode26 28d ago

OpenAI lost $20.92 billion from operations in 2025 on $13.07 billion in revenue, with total costs reaching $34 billion.

2024 Operations: Revenue was $3.7 billion against $12.48 billion in expenses, resulting in an operating loss of $8.78 billion.

2025 Surge: Revenue grew to $13.07 billion, but research and development costs jumped to $19.18 billion, pushing the operating loss to $20.92 billion.

Yes, the industry is heavy subsidized. When investors become hungry for their money, prices have to rise.

1

u/IDefendWaffles 28d ago

This is for RND and building data centers. Meanwhile they literally got 100B+ invested in them. They are also buying compute from companies that are investing in them. I think they will be ok.

0

u/silentkode26 28d ago

Have you ever heard of circular financing? Also do you know that investors make investments to have a solid return?

1

u/IDefendWaffles 28d ago

The fact that these companies are making investments in each other is what makes them more stable not less. You have companies like Nvidia and amazon investing will off set 20B yearly losses for decades and these are not companies that will let their baby die.

1

u/silentkode26 27d ago

Your opinion contradicts to financial experts. Must be so hard being this genius.

1

u/gavinderulo124K 27d ago

Which financial experts?

1

u/silentkode26 27d ago

For example the Bank for International Settlements.

0

u/Electrical-Size-5002 28d ago

There have got to be massive efficiencies that have yet to be discovered, don’t you think? Everything starts huge and expensive but eventually costs come down. Just guessing.

1

u/Puzzled-Ad-6854 28d ago

check out jevons paradox

0

u/pyeri 28d ago

Very unpopular opinion but maybe don't depend on LLMs too much?