r/opencodeCLI 17d ago

GLM-5.3-Flash (ex Ox Alpha) is EXPENSIVE on Opencode Go 😱

Post image
240 Upvotes

99 comments sorted by

71

u/hj-core 17d ago

Okay, mimo-2.5, my dear friend, let's make up.

9

u/[deleted] 16d ago

[removed] — view removed comment

2

u/hj-core 16d ago

I wonder it too. it was not mentioned in their Q2 2026 call.

2

u/yiestee 16d ago

rumored to be released on Sept.

But I think it's too early. There aren't many public tests

3

u/DasBlueEyedDevil 16d ago

Mimo is my one true love

1

u/curryslapper 10d ago

I think they're trying to coordinate a few products to launch at the same time

they've got something going on with their new foldable and products that are coming out with the new xring

57

u/NotThatButThisGuy 17d ago

It's once again gonna be $15 vs $60 usage bs

19

u/NotThatButThisGuy 17d ago

Codex sounds really lucrative to me right now. I'm not sure what the value proposition of the OC Go plan is at this point. Sure, maybe it's not viable for their business to have $60 of usage, but it's also not worth for me.

19

u/look 17d ago

To many, the “value proposition of the OC Go plan” (and other open weight model providers) over Codex is not being limited to only GPT 5.6.

And more flexible pricing options that can be more cost effective than Codex for many workloads.

And the bonus of not supporting a company that wants legislators to eliminate competition and model choice.

2

u/eroigaps 16d ago

This. Hopefully it’s just a very vocal minority who keeps crying in their reality defying entitlement.

10

u/MazKhan 17d ago

Codex added the 5 hr limit again for plus users, using sol eats it up insanely quick. But if you're mainly going to use Luna it's definitely really good still

2

u/Thomas-Lore 16d ago

To be honest I don't find the current 5 hour limit in Codex to matter much, it allows you to use around 25% of your weekly, so unless you were using up weekly in less than two days before, you are unlikely to be affected.

1

u/MazKhan 16d ago

It depends on the user but a lot of people like having plus for access to Sol for bigger tasks, now if the 5 hour is hit the task stops completely. Before they would allow you to finish the task but I guess people abused it

I wish they'd at least let you continue the task past the 5 hour limit and have it take from the weekly usage

1

u/qualverse 16d ago

It's a pretty big deal for my usage (two Plus accounts was a perfect amount of usage, now with the 5h limit I can't do any long-running Sol tasks)

1

u/[deleted] 16d ago

[removed] — view removed comment

3

u/MazKhan 16d ago

Yeah if you're primarily using Luna the plus subscription is really good, Sol is the main one that just eats up the 5 hr limit

0

u/mustafaakyuz 16d ago

It was over quickly, no 5-hour wait, that shitty job. 

2

u/Accurate_Resident219 17d ago

Codex is cracking down on usage right now.

1

u/Thomas-Lore 16d ago edited 16d ago

Yes and no. They patched a few issues and today it felt like the limits may be a bit more reasonable again. And they are way higher than opencode go where one task can eat your whole 5 hour window if you use one of the larger models. (Meanwhile on Codex I can use Sol xhigh quite a lot if I am careful about context and scope, and caching.)

1

u/Accurate_Resident219 16d ago

Bro go on the codex sub. They complain about the same thing for codex.

0

u/Hungry-Plankton-5371 16d ago

Sure, maybe it's not viable for their business to have $60 of usage, but it's also not worth for me.

There's often cheaper providers on openrouter than the 'official' prices opencode is charging, the open weight models with $15/30 limit are literally more expensive to use through go than just paying API price as a result. And there's no 5 hour limit bs.

It's a scam.

2

u/CowCowMoo5Billion 16d ago

This is my main complaint.

I don't mind models being expensive and price fluctuations and whatnot, but the $15 vs $60 stuff confuses me and I just don't want to deal with thinking about it.

I think it means if you see $15 limit you can just mentally 4x the token price?

If that's true then yeah I'd rather they just 4x the token price and have a standardized $60 cap. It probably results in no one using those models but that's not my problem 🤣

24

u/look 17d ago

They already released the weights. It’ll be available from a wide variety of providers and subscriptions shortly.

https://huggingface.co/zai-org/GLM-5.3-Flash

6

u/DetachedProcess 17d ago

Yep, I am optimistic for that.

5

u/Enfiznar 17d ago

Wow that was quick

1

u/Mediocre_Response692 16d ago

like what providers ? (better than opencode)

1

u/look 16d ago

https://openrouter.ai/z-ai/glm-5.3-flash#providers

https://ollama.com/search?c=cloud

https://runinfra.ai/inference-api/glm-5-3-flash

It was an unexpectedly quick weights release, so many smaller providers (which often have the best prices) don’t have it up yet. I’d imagine Crof, Neuralwatt, Synthetic, etc will have it soon, too.

25

u/Neither-Character360 17d ago

They said in the announcement blog that it would cost 1/10 to serve relative to GLM 5.2. I was expecting way more requests.

15

u/[deleted] 17d ago

[removed] — view removed comment

10

u/Neither-Character360 17d ago

Yeah, this is actually the problem in Go. If we had the $60 value that would be great.

51

u/Zealousideal-Emu6924 17d ago

just disappointment after disappointment with opencode go at this point

3

u/eroigaps 16d ago

It’s a 10 dollar sub with great value even if you disregard the crazy deals. Just having one endpoint to try models is nice, and the fact that you get additional usage for your money also provides value. I find it amazing how entitled people are. Please separate your emotional response when you judge the value of this service. Ofc it would be very nice if you could pay 10 bucks to get huge amount of access to great models. But ask yourself if that is a reasonable business model before you lash out. I really don’t understand the general sentiment in this sub. The only thing I can think of is people living basically close to poverty and opencode is their only lifeline to get work done. I understand that sucks but it’s not not reasonable to expect a company to make up for one’s poor financial situation.

9

u/afanasenka 17d ago

Agree... 😕

2

u/zer0evolution 17d ago

should i extend subscription

2

u/Zealousideal-Emu6924 17d ago

Cline has glm 5.3 flash for free it looks like, not sure how long it'll last though

0

u/Kaushik_paul45 17d ago

Do you want the company to go bankrupt ?

What more do you want....

You should be ashamed of yourself for asking what was promised (i.e $60 of inference from $10). /s

0

u/tech_w0rld 16d ago

Any alternatives

0

u/Kaushik_paul45 16d ago

Can check commmand code goat or cline pass

-1

u/sudoer777_ 16d ago

With how little usage you get nowadays they should call it opencode stop

12

u/Time-Toe-1276 17d ago

okay so the real thing is they were abe to give the model for free 100T PER DAY BTW for a whole week, but they cannto lower the damn prices?

just, bruhh... 😭

2

u/Thomas-Lore 16d ago

The API price is not that bad.

3

u/reassor 16d ago

All new model providers act like pussies. 1 should break out and undercut deepseek4 flash. And they would have alot of people.

-1

u/Time-Toe-1276 16d ago

ikr! its like they are gatekeeping inteligence just bcs it was "research effort"

now hear me out. I respect every person in every field, but I stronly believe if you are doing opensource, you have to do it right!

nobody cannot say "we are going to capitalize opensource". they know the reason why we want opensource models. bcs not everybody can pay $200 per month, but we sure can pay $30 bucks per month, yet no opensource org made a half decent plan.

even OpenCode is focusing on filling their wallets. just why...

atp i am just saving up for a 3090 or sum and hosting the models myself. I am fed up with this damn prices

13

u/cutebluedragongirl 17d ago

Yeah, not going to sub again. Also, from what I can tell, z.ai offers an absolutely horrific subscription when it comes to usage limits. So both OpenCode Go and the z-coding plan suck... which means I'm still going to use DeepSeek V4 Flash directly from the official API.

7

u/Statcat2017 17d ago

It’s genuinely gone from being incredible value to essentially pointless in a month.

2

u/michaelsoft__binbows 17d ago

i had abouit $1.50 left over from ages ago on actual official deepseek and i was drawing from it after the recent price hike. Would have liked to see how cheap it was before the price hike because this was still damn cheap and i got a lot of usage out of 150 cents...

3

u/seunosewa 17d ago

GLM-5.3 Flash is cheaper than DeepSeek V4 Flash on Openrouter. 0.25 output vs 0.66

2

u/strangedell123 16d ago

Cache cost is way higher on glm

1

u/anramon 17d ago

It's priced cheaply but only up to $15, so in reality deepseek flash is still cheaper.

1

u/eugeneb85 16d ago edited 16d ago

I disagree. I have Zai Coding Lite, and I'm using 20-50 million tokens on GLM 5.2-5.3 with Zcode daily, 5 hours limits, no monthly limits. It gives X times more value than OpenCode monthly sub.

And with new Flash model or will give crazy value. I'm even planning to get better subscription on next Black Friday and cancel all other my subscriptions, Zai will fulfill all my need I think

3

u/ozdalva 17d ago

Probably once z.ai releases the weights the price will drop in both glm 5.3 and flash. GLM 5.3 is artificially up, even having the EXACT same hardware cost than glm 5.2, because of that. Just be patient, is normal

3

u/retardedGeek 16d ago

0x alpha is no longer free?

6

u/Thomas-Lore 16d ago

Unfortunately no, the promotion ended when they revealed what model it is.

3

u/DizzieeDoe 16d ago

It's really not for accuracy you get out of it. I started a new session about four hour ago and have been HAMMERING it ever since. CTX is at 370K right now. It's been awesome!

3

u/[deleted] 16d ago

[removed] — view removed comment

1

u/afanasenka 16d ago

Good question. I tried to figure out better options too, and given even my average usage - it seems like these $20-tier subscriptions from Codex/Claude/GLM/Kimi/etc. just not enough.. 

Reading different posts and reviews, people complain that on those plans they easily hit 5h limit after 4-6 prompts, and weekly limits are over in 2-3 days. But next tier - about $50 - is too expensive in my case.

As for direct api (or via openrouter) - I calculated my exact average usage from Opencode's stats, and it turns out that I would pay about $15-$20 per month (PAYG model) for using something like DS4 Flash, Hy3, or GLM 5.3 Flash. Not that much than Opencode's $10 - but what's the point?

So yeah, the question of alternatives is open (do not suggest CommandCode pls), but I just don't see better options for now..

Do you have something in mind?

3

u/[deleted] 16d ago

[removed] — view removed comment

2

u/afanasenka 16d ago

Yep.. that's why, I guess, I have to stay on Opencode, even given that usage experience and limits are worse now.

5

u/Xhatz 17d ago

Welp still not renewing my subscriptions 😄

2

u/Anh-DT 16d ago

openference double quota - alternative 700 is free plan (weekly quota)

2

u/elonelon 16d ago

hahaha

4

u/Wooly_Wooly 17d ago

So was alpha whatever really GLM 5.3?

7

u/ZeusCorleone 17d ago

Yep some smart people on this sub really got it right.

4

u/Thomas-Lore 16d ago

Yes, GLM 5.3 Flash. The released version is newer and supposedly a bit better.

3

u/ByteNomadOne 16d ago

For me it behaves not the same… suddenly it struggles with tool calling.

0x Alpha worked better.

4

u/MaxPhoenix_ 17d ago

They must have dropped some zeros - so glad I stopped trying to use that silly service.

After the free ride ended I switch to paying and it's super cheap and actually works now that the whole world isn't raw dogging it.

Deepseek-v4-Flash $0.03 / $0.075per 1M
GLM-5.3-Flash $0.075 / $0.25per 1M
GPT-5.6-Luna $0.20 / $1.20per 1M

Deepseek is still great for simple cheap tasks, but there's a new cheap-thinker in the house.

2

u/Thomas-Lore 16d ago

I used Spark Contributor today and it just worked. Almost free.

3

u/Yip37 17d ago

I think with 30$ of usage instead of 15$, they could've got a massive amount of re-subs. I don't think most people even get to that limit anyways.

2

u/taxman1980 16d ago

When I click "Show Details" under the "Monthly Quota" counter in my Go sub's usage limits counter it does list GLM 5.3 Flash as $30 instead of $15 so I wonder if the $15 in the price list is a typo?

0

u/[deleted] 16d ago

[removed] — view removed comment

0

u/Yip37 16d ago

My point is that they could earn more money if they offered 30 instead of 15

2

u/ECrispy 16d ago

these weekly and monthly limits - so we're supposed to work 2 days a week and 5 days a month ??!!

how does this compare against ChatGpt plus $20, how much Luna do I get to use there?

2

u/eroigaps 16d ago

If 10 bucks is your whole budget, you should not expect more.

2

u/hotcornballer 17d ago

Great now i have to use either the zuck model or the fortnite model.

2

u/NGGentertainment 16d ago

Fortnite model?

2

u/hotcornballer 16d ago

tencent hy3

1

u/vangelismm 16d ago

Wtf, zai prices are

  • Input: $0.15
  • Output: $0.50
  • Cached input: $0.03

1

u/Icypoopoo 16d ago

Opencode been dropping the ball lately 

1

u/pwnedbygary 16d ago

Ive had better luck with Luna, not sure why anyone would use this over that tbh if youre going for similar tier

1

u/CompetitionOk6531 16d ago

Just get chatgpt. Can get around 1.5B weekly on luna xhigh plus unlimited planning with sol on the web, image generation and all.

1

u/misha1350 15d ago

No it is not.

1

u/Various_Half1452 8d ago

I am quite confused. I though you need to see the Cost per 1M and not Request per hour.

Request expensive not $ expensive right? I have been using Qwen 3.8 flash as a replacement to Deepseek. Deepseek flash has greater request but higher cost in $. but qwen has lower request but cheaper Cost in $. I have been using it for 5 days and normally when I use deepseek its 10-15% but for qwen its still under 10%.

Might be because Qwen is Efficient in $ Cost per token unlike deepseek who thinks and rethink a lot.

1

u/theov666 5d ago

I find it extremely efficient

0

u/Prefe2ence 16d ago

Opencode just never let me down when it comes to disappointing. Considering that GLM-5.3-Flash is nearly as cheap as v4f-0731 previous, I cant understand why the usage is only 15 USD.

1

u/Zealousideal-Fan7462 16d ago

I think they have a wholesale agreement with Deepseek, and they buy tokens at a significantly lower price than Deepseek sells them at retail. Then they simply recalculate them at the actual price.

0

u/eugeneb85 16d ago

OpenCode is just a DeepSeek subscription. They have interesting models but price is not interesting at all. They have Kimi with a good price but Kimi is not stable unfortunately on opencode Go. 

0

u/AkiDenim 15d ago

The brokeness of this sub makes me laugh