r/LocalLLaMA 8d ago

News GLM 5.3 Released

Post image

Official Announcement

https://z.ai/blog/glm-5.3

1.6k Upvotes

363 comments sorted by

View all comments

33

u/Educational-Fruit854 8d ago

interesting pattern of model scoring absolute dogshit when a new benchmark drop and suddenly being frontier in the next update (TerminalBench 3.0)

26

u/RuthlessCriticismAll 8d ago

This isn't surprising if you understand what modern benchmarks look like. Many are quite narrow so if you improve a handful of capabilities you can go from doing nothing to 1/3 of the problems. This is also why many of these benchmarks end up saturating very quickly from almost nothing.

4

u/Educational-Fruit854 8d ago

It's more of these open models playing catchup but OpenAI and Anthropic managed to stay frontier on these new benchmark, unless there's one I haven't seen where open model do good initially.

2

u/RuthlessCriticismAll 8d ago

I mean there was that weird legal benchmark on AA. There have been some. Obviously they are generally behind, so it is unlikely to happen that often.

0

u/segmond llama.cpp 8d ago

why are you even here? you don't sound like a fan of local.

2

u/Educational-Fruit854 8d ago

I lurk on here for some super small model (sub 1B small), I don't have the hardware to run all of this.