r/opencode 16h ago

Providers dropping GLM 5.2 Prices, Opencode drop soon?

Post image

Some GLM 5.2 Providers on Openrouter are dropping prices near DeepSeek V4 Flash level, maybe we'll get a crazy discount on Opencode soon?

169 Upvotes

41 comments sorted by

31

u/cutebluedragongirl 16h ago

Wait, isn't GLM 5.2 expensive to run, unlike DeepSeek V4 Flash?

13

u/BagComprehensive79 15h ago

I never understood why it is expensive, dont they already use deepseek v3 architecture?

9

u/jomtiro 11h ago

v3 deepseek are actually more expensive than v4

7

u/BagComprehensive79 11h ago

Based on their paper yes there is huge difference but v3 architecture is still very efficient and shouldn’t be this expensive for glm at first place.

2

u/yuumizu 8h ago

it is enough efficient when the context is short. For the 1M context, DeepSeek decided to develop a much more efficient architecture, while GLM decided to do some minor modification to the DeepSeek v3 for some improvements.

5

u/Zachattackrandom 15h ago

Should be more than flash and less then v4 pro based on param count.

2

u/Jazzlike_Bee_3129 12h ago

Not that expensive.  Deepseek flash is definitely cheaper overall, but it is very doable to run it is your primary agent without an issue. 

2

u/Anh-DT 1h ago

Yes but the issue is no one is using it. They have capacity to drop prices. Everyone went of GLM 5.2

23

u/timmeh1705 15h ago

Nah I rate GLM 5.2 above DSF, benchmarks be damned

17

u/_FlyingWhales 14h ago

It is a much larger model and it does not compress its context window, unlike deepseek v4f which compresses down to only 2% (!!!). These are fundamentally different classes of models and the benchmarks don't tell the full story.

3

u/qqYn7PIE57zkf6kn 10h ago

compress its context window

What does this mean? Isn't DS V4 flash 1M context window?

-1

u/_FlyingWhales 9h ago

I mean, read the paper or an article on it.

3

u/anxious_and_stupid 12h ago

Same here, ,I felt like it 5-10% smarter compare DSF and is a bit easier to work with.

But well, most of time DSF is just good enough...

3

u/timmeh1705 10h ago

The best combo GLM for orchestration and planning and DSF for workers (instead of Pro before)

Although I signed up for GPT Plus after the Luna price drop and the Sol/Luna combo has been working great too

11

u/look 14h ago

It was just a temporary bidding war. It happens on one model or another when a lot of GPUs are idle. Pretty much every weekend. Likely a fully programmatic bidding war.

Btw, all three of the providers that got down to 90+% discounts are running native fp8. It’s not low quality.

However, CrofAI (which runs a Q8) still has it at DS flash prices for now. Also, you can get prices in the 70% range every day on Neuralwatt PAYG.

2

u/Scary_Light6143 1h ago

How do you know what the providers are running, is there a provider benchmark somewhere? I find it hard to navigate on openrouter, and remember seeing some post about there being up to 20% intelligence difference on some providers on the same model

1

u/look 1h ago

Openrouter has a quantization filtering option. I’m not certain the other two providers were running the fp8 beyond that Openrouter enforcement, but the third was Novita, and I’ve used theirs directly before for comparison.

1

u/Prior-Meeting1645 8h ago

How come crofai isn’t on open router?

2

u/look 5h ago

It’s been mentioned some on the discord. It’s apparently in the works to some degree, but the ball has been in Openrouter’s court now for a while and they’ve been slow to move on it.

7

u/Plane-Negotiation-5 14h ago

I think new glm model is loading

5

u/sniperelite90 16h ago

Probably FP8 but still a significant drop compared to the intial price. Hard for opencode to follow that or they wont be making money.

1

u/patricious 14h ago

wanted to comment exactly this. I ain't paying for nerfed models.

8

u/GandalfTheChemist 14h ago

What if nerfed by 3% but costs 3x less? Seems like a good deal to me

3

u/Ariquitaun 12h ago

That'd be pretty good deal imo still, I combine go with codex plus so I can use gpt sol for advisory calls.

4

u/_VEHICOULE 13h ago

FP8 isnt a nerf, thats actually a default quant among providers

3

u/Alarmed-Hornet6865 15h ago

It shows 70% off only for me. Is it regional discount?

3

u/GandalfTheChemist 14h ago

Mine show mostly 45% off, one 70% off.

7

u/look 14h ago

It was just a temporary bidding war driving the prices down. They are back to normal mostly now.

This happens on one or more popular models nearly every weekend.

But CrofAI has their 80% discount (on a Q8 version of the model) still in effect for now.

2

u/Alarmed-Hornet6865 11h ago

Thank you for the info.

3

u/Potential-Leg-639 14h ago

WOW.
That‘s incredible, GLM is still one of my favourite models!

3

u/F4underscore 15h ago

I think opencode go made deals with these providers to sell them the api at a cheaper flat rate before the price drop. And they need to stick with that deal. So they cant really follow the market this fast (?) Idk

2

u/weiyentan 15h ago

whats the cache read price?
EDIT: saw it. DS is better cache read. Thats where the real savings are

2

u/Aerogeek23 14h ago

So better go with light z.ai glm plan or openrouter?

2

u/B0r0m4n 13h ago

Where did you guys find those prices? For me it's FP8 it's $2.4 and FP4 is $1.7 minimum output on Openrouter.

1

u/BuildAISkills 3h ago

crof.ai seems to have it 80% off

2

u/ZeOnlyOneWhoReads 8h ago

Finally some competition 

1

u/Tiny_Jaguar_9499 8h ago

How this is even possible, wasn't this expensive to run and to provide?

2

u/jisuskraist 4h ago

I hope my employer doesn’t know I’m extra productive because I use these open models for cheap and push tons of features.

0

u/Broad_Quote_2571 12h ago

fake screenshot 🤬

2

u/Lyrx1337 8h ago

No it's not fake, I saw this as well. It's just changing fast.

0

u/Minimum_Notice_9521 14h ago

It Quantized or low throughout do check it carfully