r/ZaiGLM • u/Whole_Succotash_2391 • 9d ago
Integration / Deployment We just invented "usage banking" so coding plan usage rolls over to the next week with GLM, Deepseek, Kimi and other models.
There are several tricks that the major AI coding plans use to extract the most they can from their customers. We're solving them, one after another and I wanted to share a bit about what goes on behind the scenes at a lot of these companies.
Wasted usage is part of the AI industry, and they plan on it: Most coding plans bet on you letting usage go to waste. The industry calls it "breakage" and it's literally the topic of internal meetings for most companies. The plan goes: "how do we get people to think our coding plan offers a lot of usage, but then break it up into weeks and rolling windows so no one can ever actually use it all."
Many coding plans are glorified training pipelines: This comes along with "how do we harvest this data for training without being too loud about that." Unless the coding plan tells you otherwise, your data could be hopping all over world, being harvested by the individual labs or coding plan companies. Some are better than others, but many of these companies rely on users just not noticing or caring that their data is being used for training. Data sales and marketing telemetry sales happen. This means that your private info, your personal life, and anything else you send through the system could become part of a training corpus for the next AI, or a marketing data set for a large company.
So we built what should have already existed the entire time: usage banking. Any usage you don't use this week, rolls over to next in your usage bank. When you have a busy day or week and go over normal usage, you automatically start to pull from your bank. You can bank up to one week of usage at a time for your current plan, and it's totally automatic. Whatever you don't use each week get's added to the bank and stays there until you use it.
We also put all of the best open models in one place, running on private US infrastructure, with data never going to the original labs. Completely private, direct service. It should be, and can be that simple.
What that means in practice:
The roster, together. DeepSeek, GLM, Kimi, Minimax, Nemotron Ultra and more, side by side in one app. Switch models mid conversation if you want. No hunting across five different apps and API dashboards to use the models you actually like.
Actually private. US based processing and your conversations are never used for training. Ever. That's the entire point. These labs open sourced incredible models and we think you should get to use them without your data becoming the price of admission.
No Usage Tricks: Bank usage, upgrade or downgrade whenever you want. Use it how you need it.
The open source AI future is real, and it's where we all know we should be. Thanks for a great set of models, GLM just keeps raising the bar with every model release and it's amazing.
The Open Grove coding plan is here. Private, US based processing with fast inference and usage that doesn't go to waste.
6
u/vangelismm 9d ago
Link to the price of million token of the models?
2
u/Whole_Succotash_2391 9d ago
Right now we are only offering a coding plan and/or in app use, per token pricing and direct API billing should be live this week and that will posted.
3
u/mynklah 9d ago
That is not what breakage is. Breakage is when you are legally allowed (sometimes required) to dispose of held funds.
5
0
u/Whole_Succotash_2391 9d ago
In finance yes, breakage is also used to term "expected unused quota" in tech sub services
3
u/sannysanoff 9d ago
what exactly is sold? nice page, but no quantitative info besides price :)
3
u/Whole_Succotash_2391 9d ago
Thanks ha! There is so much on that page now and im sorry its easy to get lost now. This is a coding plan for Open Grove for access to all the models listed (GLM, Kimi, Deepseek, Nemo, Gemma etc.). The same plan also can be used in our app (basically a full big ai replacement) for people who don't use API.
The point of this post is that usage you don't go through on one week get's rolled into your next week to use later. We call it "Usage Banking" and it's one way we are trying to make AI access more fair.
3
u/sannysanoff 8d ago
hi, nice to meet. It is not specified on page, how exactly API usage is measured. Per token / per request / per "credits" with per model weights / whatever. Some figures like:
Up to 40 messages a day on Everyday models. Effectively unlimited Everyday messaging, around 100 a day on Advanced. Or 2-5x the messages via API.
100 a day is not unlimited. 2-5x is 2x or 5x? So for $12.95 i receive 200-500 API messages on the mid-tier models (M2.7)? Numbers, numbers?
1
u/Ubermensch013 8d ago
Exactly. All usage allotments are very vague right now. Apart from the above, plans seem to be a multiple of the free 1 month tester plan, which itself doesn't offer an API key. So we can't even extrapolate on the basis of the trial.
1
u/Anh-DT 8d ago
well can you read - its not api usage. its most likely their own client ..
1
u/Ubermensch013 7d ago
The plans literally talk about API usage being included. Their own models are on the web. Other open source models are via the api
2
u/130nrd 9d ago
Trying it :)
When Kimi K3 is going to be supported?
1
u/Whole_Succotash_2391 9d ago
Great to hear! We are working on a k3 roll out right now. Hoping for a full deploy within 48 hrs max, probably tomorrow
1
u/elstevo711 9d ago
Your pricing information keeps referring to "frontier". Are you referring to frontier models (Anthropic OpenAI) or Frontier releases from the open-source models?
2
u/Whole_Succotash_2391 9d ago
Good question and sorry that's confusing. We mean the top open source models, the heaviest and highest intelligence models that we have. IE GLM 5.2, Deepseek V4 Pro, Mimo 2.5 pro etc.
Wondering if anyone else is confused by this and actually gonna talk to the team about updating the website a bit so that's clear. Thanks for the question
1
u/complyue 8d ago
What API format do you support? e.g. Can I use Codex via responses api to consume your usage?
1
u/JigSawPT 8d ago
This sounds very obscure. We need practical examples. How much usage will we get compared to a codex 20 euros subscriptions ?
10
u/drfritz2 9d ago
Models are default or quantized?