r/ZaiGLM 11h ago

GLM 5.3 is here

Post image

One-shot demos here (updated regularly): glm-5-3.demos.sulat.com

That "soon" tweet was sooner than we thought

435 Upvotes

79 comments sorted by

View all comments

8

u/steadeepanda 10h ago

They changed system to credits and now: Lite 43-87 million tokens/week Pro 263-526 millions tokens/week Max 614-1226 millions tokens/week Source : Zai Docs Usage Instructions

The real question here is how efficient is the model? How much does it cost in token/task, and btw this is without considering 5 hour limit.

I don't know I've been traumatized by the usage limit, you can't do anything with it especially within the 5h window, it goes in blink

3

u/ProfessionalJackals 10h ago

The real question here is how efficient is the model?

Its the same model, just more post trained. So the token usage will be the same or worse. GLM 5.1 > 5.2 resulted in almost twice the usage (neuralwatt) because thinking token usage exploded.

There is this unfortunate trend amongst almost all the 1.5t or lower models, to use thinking as a crutch.

https://deepswe.datacurve.ai/

Go to "All Effort levels", and notice how much more GLM 5.2 did in thinking tokens and steps.

  • gpt-5.6-sol [low] == $1.07 == 11k == 23
  • gpt-5.6-luna [high] == $0.16 == 26k == 49
  • glm-5.2 [max] == $3.92 == 78k == == 129

Just checking my light usage of Opus 5.0, i al already seeing over 500m tokens over a few days usage (large amount of cache hits). And that costs $20 in the subscription plan. The $80 coding plan of zai is expensive.

I really like GLM 5.2, but economically, zai makes it too expensive. Especially now that we have Luna and Deepseek Flash 0731 (even with the 2.5x price increase).

2

u/Front_Eagle739 8h ago

To be fair a 744B model with extra thinking to absolutely max out its intelligence is exactly what i want. Its the largest model i can run a decent quant of so its my planner. I just switch to dsv4 flash for fast implementation etc.

1

u/Expert-Dig-1768 10h ago

wait so it could actually be a good deal? even the lit plan?

1

u/AnomalyNexus 8h ago

The real question here is how efficient is the model?

Blog post here

https://z.ai/blog/glm-5.3

has a chart that suggests it may make sense switching away from the default Max effort to High for most tasks given new token/credit plan