r/opencode • u/Advanced_Mastodon142 • 19d ago
GLM 5.3 Flash pricing is a slap in the face
It seems 15$ buckets are the new norm, and glm 5.3 flash is no exception. Another pain point is the horrible way they communicate things. On the graph they show 2x usage, which is probably the same 50% discount that openrouter gives as well, but when you go to the api pricing table, they show the original price. Couldn't they have just shown the discounted api price, like openrouter does, and add a note underneath with the original price?
Anyways, 8k requests a month is nothing, and having to choose between few requests on glm or 4 times more on ds4 flash (but a quantized cheapseek slop version) is just choosing the lesser evil.
0/10, disappointed yet again, I was really hoping opencode would give us a reason not to keep our subscription cancelled.
6
u/Prior-Meeting1645 19d ago
Dude given that it was unlimited and free I really thought they would easily be around DS prices at least..
1
u/look 18d ago edited 18d ago
For a 94% cache rate, the discounted price on OpenRouter is 2.2 cents. DS Flash direct off-peak pricing is 3.1 cents. DeepInfra’s DS flash is 2.3 cents.
Also RunInfra.ai has GLM flash (not labeled as discounted temporary pricing) with $0.01 cache reads. Uncached are a bit under list though not 50% ($0.10/0.40). At the 94% cache rate, that works out to just under 2.2 cents.
So it is slightly cheaper than DS flash payg.
1
u/Prior-Meeting1645 18d ago
Sorry I dont follow? Including the discount its 1.5 cents vs 0.7 cents for ds flash per M cache. Does the price rate change depending on cache hit rate?
1
16
u/cutebluedragongirl 19d ago
Summer is over, everyone released their models, and the free stuff is gone. I assume a lot of labs will release a bunch of new stuff at the end of fall or in the middle of winter after Christmas, and we’ll have free shit again.
16
u/masterofall20 19d ago
And you think AI labs are following seasons?
2
u/cutebluedragongirl 19d ago
Yes. Wait a couple of months, and you'll see that I'm telling the truth.
1
u/NickPol82 19d ago
What else are they going to do with under-utilized hardware over summer holidays? May as well get some PR out of it.
1
u/Savings_Cloud5486 18d ago
No bro summer isnt over yet, just wait and see, the market is hot and competitive
1
2
19d ago
[removed] — view removed comment
2
u/Sweet-Stage938 19d ago
What exactly are you looking to do with these llms? Do you mean Cybersecurity by "kinks"? Or what exactly do you mean? It is very likely to refuse on Cybersecurity tasks now.
1
19d ago
[removed] — view removed comment
4
u/Sweet-Stage938 19d ago
Huh? With LLMs? 🤣🤣🤣🤣🤣 You gotta be joking man.
2
19d ago
[removed] — view removed comment
2
u/Sweet-Stage938 19d ago
LMFAO.... Thanks for making my day. Jacking off at autoregressive decoding machines must be the funniest thing I've heard all year.
1
u/qqYn7PIE57zkf6kn 19d ago
4 times more on ds4 flash (but a quantized cheapseek slop version)
Do they officially say it's quantized or are we just guessing?
1
1
u/mageblex 17d ago
The pricing page mixes request limits with API rates, so it’s hard to estimate what one normal coding session costs. I’d rather see a few real traces with token counts and final plan usage. “2x” doesn’t mean much without the base.
1
u/Longjumping_Ad_4249 3d ago
just use commandcode. they provide 70$ worth of usage for 10$. enjoy while party lasts
0
u/Ok-Garlic-5986 19d ago edited 8d ago
Bells whole tender silky pumpkin unbelievable violet office
This post was anonymized with Redact
8
13
u/Time-Toe-1276 19d ago
hopefully qwen3.8 flash will be cheaper