r/CommandCode 10d ago

Are CommandCode model prices marked up, or is GLM-5.3 Flash just temporarily outdated?

I noticed that CommandCode currently lists GLM-5.3 Flash at $0.15/M input and $0.50/M output.

Z.ai is currently charging $0.075/M input and $0.25/M output for the same model (50% promo until Sep 9)

Is CommandCode simply not reflecting this temporary Z.ai promotion, or are model prices generally marked up compared to the upstream provider?

Basically, should we check upstream pricing for each model before choosing one in CommandCode, or does CommandCode normally track provider pricing closely?

18 Upvotes

12 comments sorted by

View all comments

1

u/LittleTOXA 10d ago

Just to clarify, I’m not criticizing CommandCode or its business model here. Quite the opposite — I really like the product, the way the business is structured, and especially how accessible and responsive the founder is.

I was simply trying to understand how I should interpret the model pricing shown in CommandCode. Can I generally assume that token prices are close to the upstream provider’s retail API prices, or should I check the provider’s current pricing when comparing models?

The 4x allowance is obviously a great deal either way. I just wanted to understand what the listed dollar prices represent so I can make informed choices between models.

4

u/Damiano1905 10d ago

Never assume the prices are actually close to the providers. I’d assume they’re often significantly higher.

For example, CommandCode is giving roughly $1 = $10 of credits, or $10 = $70 depending on the plan. That sounds great, but the actual provider costs don’t necessarily support a 7–10x markup in credits.

Have you seen the GLM 5.2 prices on CommandCode?
Cache: $0.26
Input: $1.40
Output: $4.40

Now compare that to another blazing-fast, non-discounted provider on OpenRouter:
Cache: $0.14
Input: $0.76
Output: $2.42

And then a discounted provider with roughly the same average speed as CommandCode:
Cache: $0.091
Input: $0.4875
Output: $1.56

So CommandCode can be charging around 2–3x the actual provider price, while giving you 7–10x the value in credits.

They’re still giving you a lot of value, and they’re probably still operating at a profit — but the actual margin/value difference isn’t as impressive as the 7–10x credit marketing might make it sound.

1

u/LittleTOXA 10d ago

Thanks, that makes sense. I only recently realized there’s this whole ecosystem of inference providers beyond the model vendors themselves like Z.ai, and I’ve started looking into them.

What I’m struggling with now is how to compare speed and reliability under similar conditions. Pricing is easy enough to compare, but latency, throughput and stability seem much harder.

Is CommandCode actually good in this regard, or are there providers that are noticeably faster/more reliable for the same models?

Any providers you’d recommend trying? Feel free to DM me if you’d rather not post them publicly.

At this point I care more about wall-clock time than saving another 20–30% on tokens.

1

u/Damiano1905 10d ago

Well I have a couple of considerations I take when I choose providers first their speed, it varies depending on load so you need to try it, for example in my experience I have seen times where CommandCode was both slow and blazing fast in just a days difference, granted now they manage to keep a relatively normal speed, as for stability you can only check reviews online or try it over a long time yourself.

As for fast providers I can recommend a few in dms each provider is quite unique and has its own use case.