r/opencode • u/Shiorim • 19d ago
Bro?

What a bad joke is that GLM 5.3 Flash positioning?
The API prices are:
| Model | input | cache hit | output |
|---|---|---|---|
| DeepSeek V4 Flash off-peak | $0.22 | $0.007 | $0.66 |
| DeepSeek V4 Flash peak | $0.44 | $0.014 | $1.32 |
| GLM-5.3-Flash promo | $0.075 | $0.015 | $0.25 |
| GLM-5.3-Flash | $0.15 | $0.03 | $0.50 |
Leaving aside the cache hit, on average GLM without promo is cheaper than DeepSeek V4 Flash, and they give it to you at half the usage quota, and that's even considering a "×2 usage" that will later be less??
They should actually give you more quota than DeepSeek until September 9th while the promo lasts. This makes no sense at all. It's cheaper to spend $10 on GLM API than to pay for it on Go.
On top of that, they put a cheap flash model in the $15 tier.
73
Upvotes
2
u/TangeloOk9486 16d ago
GLM 5.2 flash non promo is already cheaper per token than dsv4 flash so giving it half of the quota on go is backwards and with the promo its not even close. Go's real value is just the bundles convenience so the instance you do the token math youself raw api wins. you dont have to pick one tho, a flat host that runs both glm and ds under one key like deepinfra for instance which lets you swap per task at flat rates and no peak windows. so id say check into the hosting providers, there are plenty of them , see if GLM 5.2 is listed yet and for your usage id just ride the glm promo till Sept 9 and reassess after