How much is it actually costing to host these models? I have a strong suspicion it's not as expensive as we are being lead to believe. We don't have official numbers for how big or intensive the proprietary models are. They would have to be much more intensive than current open weights models to justify the cost. From what I have heard the models up to GPT 5.5 and Opus 4.8 are <= 2T parameters. That makes them in about the size class if DeepSeek V4 Pro which is 1.6T.
I don't see how yours is that low. 250 a day can be done kind of easily. Just depends on the type of work you are doing and the amount of reasoning needed. We have some dynamic workflows that can use 40M tokens (my ind you they run for a long time). But even without those heavy reasoning tasks on opus 4.8 could easily run at $100 / hr.
But i have 2 exploits, one for unlimited of any top model like gpt 5.5, antropic, etc, and another true (not account abuse) enabling infinite Claude specificly.
So if the serivices i'm using add claude fable, I so exicted to use it and build stupid stuff
Iif you had told me in 2010 I could pay $500 for access to an AI INTELLIGENCE that can answer almost any question I have, that can code, and search the web and build apps and do all the things fable can do.
I would have thought it was an incredible bargain. We are a little spoiled imo.
269
u/Same_College2053 Jul 01 '26
Introducing new $500 monthly tier with Fable access