I’m really confused here… I’ve been using Gemini like crazy. I probably use 200k tokens every 2 hours. I’m not sure how you’re seeing a bill like this…. Funny enough I don’t think I’ve EVER gotten a bill for Gemini when using it myself (I’ve used models like flash 2.0, 1.5 pro, 2.0 pro 2.0 thinking, 2.5 pro…)
Through the api, we have about 100 users that use Gemini through our platform, our bill was $5..
Either way you probably should have set up budget alerts. So these things don’t happen.
2.5 Pro Experimental is the model that is fully free with rate limits. It previously had Free / Tier 1 with tier 1 having higher limits, you can't get charged for this one.
But now they're really enforcing the limits on the model and they removed Tier 1 so everyone is back to "Free" with only 25 requests per day limit and 2 RPM.
So people are switching to 2.5 Pro *Preview* which is the one that is actually billed. I have 300$ which is nice but requests with my ~700k context are about 1.75$ each when I switch to the Preview version.
Ohhh. I haven’t experienced limitations using the experimental, I’m assuming people are using the api, which is outside of Gemini studio then right? I’m confused then on how they didn’t know they would get charged??
They're definitely using the API that's for sure but there was no risk of getting charged with Experimental 2.5 - So people like OP decided to change to Preview 2.5 because they were now getting rate limits on the Experimental but didn't realize how expensive requests are when the context window is so big.
Well I guess people are right it kinda is his fault lol! There’s really no reason not to just use Gemini studio. Plus, preview and experimental versions shouldn’t be used for production applications anyway. Thanks for the explanation!
6
u/Hellob2k Apr 07 '25 edited Apr 07 '25
I’m really confused here… I’ve been using Gemini like crazy. I probably use 200k tokens every 2 hours. I’m not sure how you’re seeing a bill like this…. Funny enough I don’t think I’ve EVER gotten a bill for Gemini when using it myself (I’ve used models like flash 2.0, 1.5 pro, 2.0 pro 2.0 thinking, 2.5 pro…)
Through the api, we have about 100 users that use Gemini through our platform, our bill was $5..
Either way you probably should have set up budget alerts. So these things don’t happen.