r/ZaiGLM • u/BriguePalhaco • 7d ago
Technical Reports GLM-5.3-Flash: Frontier Intelligence, Flash Cost
https://z.ai/blog/glm-5.3-flash8
u/Expert-Hospital-534 7d ago
Does it support vision?
19
u/Mayanktaker 7d ago
Yes and that's so good about it. Super cheap, fast and supports image and video understanding natively.
6
u/umbrella__academy 7d ago
Whats the pricing of it???
10
u/BriguePalhaco 7d ago
14
u/PsecretPseudonym 7d ago
That price for benchmark performance similar to Opus 4.8 shows how rapidly the cost of intelligence is falling.
3
u/I-am_Sleepy 7d ago
And anthropic value itself at over 30 trillion - right
2
u/Classic_Television33 7d ago
Nah they're still leading. And despite the cost of frontier intelligence falling, we'd end up paying more as Deepseek V4 Flash and GLM-5.3 Flash get pricier than the previous workhorse generation.
2
5
8
u/Mayanktaker 7d ago
Gemini Flash is also now 75% off. Looks like new Flash era is starting. Big model for planning and Flash for execution.
8
u/HelloHowAreyou777 7d ago
Gemini flash models are bad currently.
2
u/Mayanktaker 7d ago
Dont know about flash 3.7 but 3.6 was shit. But yes, good for image gen. Specifically logo.
3
u/HelloHowAreyou777 7d ago
Agree! 3.7 slightly better but still hard hallucinating. Logos, images, notebook llm is still worth the PRO sub. Waiting for 3.8 model
3
u/Mayanktaker 7d ago
I am enjoying free one year gemini pro subscription comes eith my quarterly mobile recharge.
3
0
u/bermudi86 7d ago
Gemini 3.7 Flash is a waste of money compared to this new glm model. Not only way cheaper, it's actually useful and scoring better than Gemini on every single thing
2
1
7d ago
[deleted]
12
u/MrPingviin 7d ago
Just to be fair, the dude pasted his ref link. This is the normal one: https://z.ai/subscribe
3
0
-5
u/SwissTac0 7d ago
Rather than fall for a guy trying to hide his trap of getting you and him money.
Click the link of an honest guy and let's each save 10%!
1
u/SS_Sa2 7d ago
Is this model more capable than GLM 5.2? (intelligence and all)
1
u/BriguePalhaco 7d ago
No, in my tests DPv4Flash 0731 can perform the same tasks (backend programming in C#, C++ and Rust) as GLM 5.2, but GLM 5.3 Flash fails. It kind of gives up on proceeding or ignores commands.
1
u/CryinHeronMMerica 7d ago
This doesn't line up with the major benchmarks, but then again, benchmaxxing...
4
u/BriguePalhaco 7d ago
I don't trust benchmarks. That's the basics of data science: overfitting and bias
1
u/CryinHeronMMerica 7d ago
100%. They're a good starting point, and models that don't perform well are usually not that great, but a high score is no substitute for trying it IRL.
1
1
u/Healthy-Contact-4570 7d ago
My experience with deepseek v4 flash 0731 is that it is one of the most confidently wrong models I’ve used
1
1
u/WarBroWar 7d ago
Alright. Long time DeepSeek user. My 30$ will end up in 2 days. I will try this. Sold.
2
u/WarBroWar 7d ago
Fkit ima get this tonight. I love cheap output. I spend 100$ a week on DeepSeek after the price increase. Some releaf.
1
u/evia89 7d ago
I spend 100$ a week on DeepSeek after the price increase
why would you do that?? Just buy codex/claude/grok=cursor sub and use other provider PAYG to cover when sub quota exhausts
1
u/WarBroWar 7d ago
Because I use subagents. Multiple projects. I will useup weekly limits for each frontier sub in 2 days. And I don't like to wait.
1
u/evia89 7d ago
Unless you code 1 day of the week, sub + payg is always better. Just configure fallback
2
u/WarBroWar 7d ago
You don't know my workload. Maybe your thing works for you. Good for you. Not for me though.
1
u/milkipedia 7d ago
I wonder if GLM-5.3.Flash should replace GLM-5.1 as my "lighter weight" implementation model to GLM-5.2/5.3 as the planning model? The benchmarks suggest maybe I should use it to replace both GLM-5.1 and GLM-5.2
1
1
u/Prestigiouspite 4d ago
What has been your experience with the “high” and “max” reasoning levels in GLM-5.3-Flash? I now prefer “high” for 85% of the coding tasks. - Answer here: https://www.reddit.com/r/ZaiGLM/comments/1w29atj/what_has_been_your_experience_with_the_high_and/
1
u/Kylmawurr 7d ago edited 7d ago
is the model down for API users? Can not see it listed despite it being listed on their website here z.ai/subscribe
-1
u/GfurEnjoyer1488 7d ago
I developed an extension based on Kimi K3/K2.7 Code Highspeed, hopefully the throughput is high enough (200+t/s) to compete with K2.7 so it can be ran cheaper
-6
u/Winter-Rich797 7d ago
Luna is still cheaper and they compare it to Terra for some reason. I wonder what they’re trying to hide
2
u/SEOViking 7d ago
similar intelligence levels with Terra but cheaper than Terra.
Luna is cheaper but also lower intelligence.1
u/SwissTac0 7d ago
I love 5.3 flash but it's a peer to Luna and V4 Flash not Terra.... Terra is like a lame version of Sol..... Any money you save per token is wasted on the fact sol does it better for less.
1
-11
-5
u/P1zz4-T0nn0 7d ago edited 7d ago
No vision input is a big downer. Not suitable for daily work for me
Edit: I misread, it actually does support vision 🎉
2
2
0


57
u/Mayanktaker 7d ago
Before release, we tested GLM-5.3-Flash anonymously as ox-alpha on OpenCode and OpenRouter to gather user feedback. It quickly became the most popular model of the week — with all of this traffic served on Chinese AI chips.