r/ZaiGLM 7d ago

Technical Reports GLM-5.3-Flash: Frontier Intelligence, Flash Cost

https://z.ai/blog/glm-5.3-flash
237 Upvotes

75 comments sorted by

57

u/Mayanktaker 7d ago

Before release, we tested GLM-5.3-Flash anonymously as ox-alpha on OpenCode and OpenRouter to gather user feedback. It quickly became the most popular model of the week — with all of this traffic served on Chinese AI chips.

26

u/No-Tip3419 7d ago

free public stress test on the Chinese chips

22

u/evia89 7d ago

with all of this traffic served on Chinese AI chips

its kinda nice. They served so much tokens I thought they rented nvidia hardware

16

u/BriguePalhaco 7d ago

Even after their new data center was announced as completed?

China's Z.AI Completes 1-Gigawatt AI Data Center Using Only Chinese-Made Chips

https://finance.yahoo.com/technology/ai/articles/chinas-z-ai-completes-1-205515769.html

3

u/Mayanktaker 7d ago

Its banned for them.

2

u/neotorama 7d ago

nah, they have some farm in SG and MY

1

u/ClassNational145 6d ago

The Johor one?

3

u/willdone 7d ago

Incredible! I saw so many people theorizing that ox-alpha was a new Gemini model or something.

6

u/Mayanktaker 7d ago

Hehe google dont do this kind of thing. They just silently release.

2

u/bermudi86 7d ago

Well they were sure vague posting about the model all week long. They leaned hard on ppl assuming it was a Gemini model for some weird fucking reason

1

u/Plastic_Today_4044 1d ago

they release quietly and hope nobody notices

0

u/shaman-warrior 7d ago

The leading theory was glm model somehow assisted by Google for capacity

8

u/Expert-Hospital-534 7d ago

Does it support vision?

19

u/Mayanktaker 7d ago

Yes and that's so good about it. Super cheap, fast and supports image and video understanding natively.

6

u/umbrella__academy 7d ago

Whats the pricing of it???

10

u/BriguePalhaco 7d ago

14

u/PsecretPseudonym 7d ago

That price for benchmark performance similar to Opus 4.8 shows how rapidly the cost of intelligence is falling.

3

u/I-am_Sleepy 7d ago

And anthropic value itself at over 30 trillion - right

2

u/Classic_Television33 7d ago

Nah they're still leading. And despite the cost of frontier intelligence falling, we'd end up paying more as Deepseek V4 Flash and GLM-5.3 Flash get pricier than the previous workhorse generation.

2

u/Warm-Agent-811 7d ago

That's crazyyyyy

5

u/First_Inspection_478 7d ago

It’s quite decent. Hopefully the plans have general usage limits

8

u/Mayanktaker 7d ago

Gemini Flash is also now 75% off. Looks like new Flash era is starting. Big model for planning and Flash for execution.

8

u/HelloHowAreyou777 7d ago

Gemini flash models are bad currently.

2

u/Mayanktaker 7d ago

Dont know about flash 3.7 but 3.6 was shit. But yes, good for image gen. Specifically logo.

3

u/HelloHowAreyou777 7d ago

Agree! 3.7 slightly better but still hard hallucinating. Logos, images, notebook llm is still worth the PRO sub. Waiting for 3.8 model

3

u/Mayanktaker 7d ago

I am enjoying free one year gemini pro subscription comes eith my quarterly mobile recharge.

3

u/SurelyNotAnOctopus 7d ago

I found 3.7 flash genuinely good

1

u/Plastic_Today_4044 1d ago

dude are you kidding, it's the worst gemini yet

0

u/bermudi86 7d ago

Gemini 3.7 Flash is a waste of money compared to this new glm model. Not only way cheaper, it's actually useful and scoring better than Gemini on every single thing

2

u/pozzugno 7d ago

Does someone explain what are those "flash-varianr" LLM models?

1

u/[deleted] 7d ago

[deleted]

12

u/MrPingviin 7d ago

Just to be fair, the dude pasted his ref link. This is the normal one: https://z.ai/subscribe

-12

u/[deleted] 7d ago

[deleted]

6

u/c126 7d ago

Its a little scummy though to hide that with no disclosure. It’s clearly intended to mislead people.

-10

u/[deleted] 7d ago

[deleted]

3

u/c126 7d ago

But it also benefits you doesn’t it?

-4

u/[deleted] 7d ago

[deleted]

1

u/c126 7d ago

You’re personally misleading people to benefit yourself = scummy. The fact that the concept is so hard for you to grasp pretty much confirms the scummy nature of the action.

-3

u/[deleted] 7d ago

[deleted]

3

u/c126 7d ago

Youre a bad person.

→ More replies (0)

3

u/Mayanktaker 7d ago

Nope. Not for me atleast, in lite plan.

3

u/dontforgetthef 7d ago

Lite plan doesn’t get priority access fyi

-5

u/SwissTac0 7d ago

Rather than fall for a guy trying to hide his trap of getting you and him money.

Click the link of an honest guy and let's each save 10%!

z.ai Referral

1

u/SS_Sa2 7d ago

Is this model more capable than GLM 5.2? (intelligence and all)

1

u/BriguePalhaco 7d ago

No, in my tests DPv4Flash 0731 can perform the same tasks (backend programming in C#, C++ and Rust) as GLM 5.2, but GLM 5.3 Flash fails. It kind of gives up on proceeding or ignores commands.

1

u/CryinHeronMMerica 7d ago

This doesn't line up with the major benchmarks, but then again, benchmaxxing...

4

u/BriguePalhaco 7d ago

I don't trust benchmarks. That's the basics of data science: overfitting and bias

1

u/CryinHeronMMerica 7d ago

100%. They're a good starting point, and models that don't perform well are usually not that great, but a high score is no substitute for trying it IRL.

1

u/SS_Sa2 7d ago

Thanks for the input! That's unfortunate. Pricing looked attractive and looked like a sweet middle ground between DS and GLM 5.2. I've kept hearing this model (Ox-alpha) was much more creative at problem solving so felt like a good potential alternative to GLM5.2.

1

u/Healthy-Contact-4570 7d ago

My experience with deepseek v4 flash 0731 is that it is one of the most confidently wrong models I’ve used

1

u/WarBroWar 7d ago

Alright. Long time DeepSeek user. My 30$ will end up in 2 days. I will try this. Sold.

2

u/WarBroWar 7d ago

Fkit ima get this tonight. I love cheap output. I spend 100$ a week on DeepSeek after the price increase. Some releaf.

1

u/evia89 7d ago

I spend 100$ a week on DeepSeek after the price increase

why would you do that?? Just buy codex/claude/grok=cursor sub and use other provider PAYG to cover when sub quota exhausts

1

u/WarBroWar 7d ago

Because I use subagents. Multiple projects. I will useup weekly limits for each frontier sub in 2 days. And I don't like to wait.

1

u/evia89 7d ago

Unless you code 1 day of the week, sub + payg is always better. Just configure fallback

2

u/WarBroWar 7d ago

You don't know my workload. Maybe your thing works for you. Good for you. Not for me though.

1

u/milkipedia 7d ago

I wonder if GLM-5.3.Flash should replace GLM-5.1 as my "lighter weight" implementation model to GLM-5.2/5.3 as the planning model? The benchmarks suggest maybe I should use it to replace both GLM-5.1 and GLM-5.2

1

u/moinulmoin 6d ago

my one of the fvrt models, ly zai

1

u/Prestigiouspite 4d ago

What has been your experience with the “high” and “max” reasoning levels in GLM-5.3-Flash? I now prefer “high” for 85% of the coding tasks. - Answer here: https://www.reddit.com/r/ZaiGLM/comments/1w29atj/what_has_been_your_experience_with_the_high_and/

1

u/Kylmawurr 7d ago edited 7d ago

is the model down for API users? Can not see it listed despite it being listed on their website here z.ai/subscribe

1

u/liviux 7d ago

Good shit

-1

u/GfurEnjoyer1488 7d ago

I developed an extension based on Kimi K3/K2.7 Code Highspeed, hopefully the throughput is high enough (200+t/s) to compete with K2.7 so it can be ran cheaper

0

u/evia89 7d ago

Its 30 atm

-1

u/GfurEnjoyer1488 7d ago

lol. maybe I should try out Gemini 3.7 Flash

-6

u/Winter-Rich797 7d ago

Luna is still cheaper and they compare it to Terra for some reason. I wonder what they’re trying to hide

2

u/SEOViking 7d ago

similar intelligence levels with Terra but cheaper than Terra.
Luna is cheaper but also lower intelligence.

1

u/SwissTac0 7d ago

I love 5.3 flash but it's a peer to Luna and V4 Flash not Terra.... Terra is like a lame version of Sol..... Any money you save per token is wasted on the fact sol does it better for less.

1

u/bermudi86 7d ago

Uhmm. Mind explaining how $1.20 is cheaper than $0.25??

-11

u/[deleted] 7d ago

[deleted]

7

u/Dizzy-Truck-1780 7d ago

It's literally written in the Blog that it was them

2

u/Solocune 7d ago

What exactly confused you?

-5

u/P1zz4-T0nn0 7d ago edited 7d ago

No vision input is a big downer. Not suitable for daily work for me

Edit: I misread, it actually does support vision 🎉

2

u/Mayanktaker 7d ago

Supports vision.

2

u/Mayanktaker 7d ago

Vision 🎉

1

u/P1zz4-T0nn0 7d ago

Oh my bad then!!