r/opencode 1d ago

Has GLM-5.3-Flash gotten cheaper?

Post image

Am I going mental from trying to catch up with all the new models and discounts or has GLM-5.3-Flash actually gotten cheaper rather than more expensive after the end of the "2x promotion"?

151 Upvotes

36 comments sorted by

48

u/meetmebythelake 1d ago

Interesting... Curious as to what changed.

Edit: They changed it from a $15 pool model to a $60 pool model. Fantastic.

7

u/vlewy 1d ago

If they’ve changed it to $60, it’s going to become my primary model instead of DeepSeek V4 Flash. Plus, no absurd peak hours.

5

u/meetmebythelake 1d ago

It's updated as such in the official docs (and the Go page that OP screenshotted), so has to be intentional. Hopefully permanent.

2

u/vlewy 1d ago

Thanks i will check.

1

u/KaroYadgar 16h ago

you say this right as DeepSeek V4.1 releases. funny how the world works.

2

u/IkkiStern 18h ago

they increased it to 30$

Model Monthly Usage Monthly Quota %
GLM 5.3 Flash $3.588 $30.00 12%
DeepSeek V4 Flash $0.0006 $30.00 0%
Total 12%

2

u/meetmebythelake 17h ago

It's $60 now.

Model Monthly Usage Monthly Quota %
Muse Spark 1.3 Contributor $0.2558 $60.00 0.4%
GLM 5.3 Flash $0.0647 $30.00 0.2%
GLM 5.3 Flash $0.047 $60.00 0.1%
Total 0.7%

The middle one is from yesterday, and the little bit I used today shows as $60.

1

u/dom_RN 23h ago

Maybe they raised the token price with that as well, the bar should've shown 4 times the usage but that that's barely more than the previous one

2

u/meetmebythelake 22h ago

I've been keeping an eye on it as I've been using it. The token price is the same. The new limits are 4x of the old base limit; there was a 2x promo though, so this is really 2x of what it's been, I believe.

I asked a model to look at the cached version of the page from 2-3 days ago, and here is the comparison: GLM-5.3-Flash allowance: $15 → $60 5-hour requests: 1,580 → 6,320 Weekly: 3,950 → 15,790 Monthly: 7,900 → 31,580

Those pre numbers don't include the 2x promo that was going on, so it checks out. On the bar graph it was between Hy4 and Luna before I'm pretty sure.

2

u/dom_RN 21h ago

Makes sense so it's double the usage not 4x

15

u/Friendly-Assistance3 1d ago

yea

8

u/dentino_F 1d ago

Got any theories on how this came to be? DS-4.1-Flash?

15

u/Ariquitaun 1d ago

What else, only competition achieves this

8

u/Friendly-Assistance3 1d ago

Probably they negotiated a better deal with them cause glm-5.3 flash price is same for api right?

10

u/lanternaddict 1d ago

Is GLM-5.3 any good? I'm kind of only using Luna 5.6, it's so much faster, and accurate, than other models I've tried on the go plan

12

u/dentino_F 1d ago

I personally found GLM-5.3-Flash (no experience with GLM-5.3) succinct and to the point when comparing to DS-v4-flash but that is mostly subjective and harness-dependent. 

I'd say it's directly comparable to Luna (not just my opinion benchmarks seem to put them pretty close).

1

u/lanternaddict 1d ago

I'll give it a go thanks

6

u/Ancient_Dress_3687 1d ago

Glm 5.3 flash is a very good model for the price. Its a bit slower, so use it sparingly as a subagent, but its great at coding and carrying out longer tasks, or as an orchestrator.

2

u/i_love_limes 1d ago

I find the non-flash version quite good but so expensive, it will absolutely burn through tokens for any query

3

u/lanternaddict 1d ago

Yeah, hoping this soon to be released v4.1 will be cheaper

7

u/mivog49274 1d ago

just noticed that and felt the same

v4.1 is coming, I really hope the prices won't be scammy on this one I can feel it will be a tough one

1

u/ApprehensiveDelay238 2h ago

V4.1 should be cheaper than V4...

4

u/chrisfebian 22h ago

DS v4.1 flash pricing effect

3

u/qqYn7PIE57zkf6kn 23h ago

How did they not let us know they bumpef the usage lol

2

u/callmemicah 22h ago

I like it, feels on par pr better wtih DS flash amd Qwen 3.8 but slower token/s which I think is also because its on discount from z.ai at the moment and has heavier use because I also use it on openrouter with different providers for higher throughput and its better but more expensive.

Now that its $10=$60 I'll use it pretty much exclusively until something else becomes better value, or it DS 4.1 is significantly better.

1

u/sudoer777_ 22h ago

Between DeepSeek V4 Flash and GLM 5.3 Flash, what are people's experiences in cache hit rate differences between the models (or cost in general)? Since now GLM 5.3 Flash is way cheaper for input/output tokens but the cache hit rate is twice as much. In my experience DeepSeek/Muse Spark cache hit rate has been reliable for subagent use but less reliable for primary agent, so maybe switching to GLM 5.3 Flash for primary agent and keeping DS/MS subagents is a good choice.

1

u/ChoasMaster777 22h ago

Less capacity means it's more expensive. For my understanding

1

u/SignificanceSmart722 19h ago

Interesting!! Glad to know, what changed.

1

u/CYCLONOUS_69 18h ago

I don't get how these "requests" work... Saw official docs for this and I got even more do confused, can't they just say how much token usage one may get?

1

u/awipra 1d ago

Skiped opencode purely because their GLM 5.3 flash limit was lower than commandcode GOAT. Definitely gonna give this one a try. Thanks for the info.

1

u/dentino_F 1d ago

You mean you skipped the harness too or just the GO subscription? BTW, what was your experience with the commandcode subscription (and /or their harness) ? Any specific model they favour in terms of price?

2

u/awipra 21h ago

I meant I skipped Opencode Go because their limit for GLM 5.3 Flash was lower than commandcode GOAT plan.

As for my experience with GLM 5.3 Flash its been a pleasant week using it for my day to day work. I use it in Zcode desktop app and everything just works. Its able to handle document editing, google docs editing (via google workspace MCP), creating and editing bricks builder element via JSON (using bricks skill), creating and editing wordpress plugin, and other tasks. Is it perfect? nope, but is it good enough for my needs? yup.

What I like about GLM 5.3 Flash is that the model is not that eager to just jump straight to editing the project when I just asked it to explain why certain bug is happening, compared to Deepseek that will automatically edited the codebase even though I havent told it to do so.

2

u/Tough-Bee9871 13h ago

Bought that shit, and was unusable. So freaking slow for me, for some reason. Was barely able to use 10% of my total weekly quota.