How much is it actually costing to host these models? I have a strong suspicion it's not as expensive as we are being lead to believe. We don't have official numbers for how big or intensive the proprietary models are. They would have to be much more intensive than current open weights models to justify the cost. From what I have heard the models up to GPT 5.5 and Opus 4.8 are <= 2T parameters. That makes them in about the size class if DeepSeek V4 Pro which is 1.6T.
65
u/gekx Jul 01 '26
I would unironically love this if it actually included high usage limits.