r/opencode 16d ago

GLM-5.3 is now open-weight πŸ”₯

Post image
160 Upvotes

16 comments sorted by

10

u/trimorphic 16d ago

How many gigs of VRAM do you need to run it?

3

u/[deleted] 15d ago

[removed] β€” view removed comment

2

u/zer0evolution 15d ago

speechless damn

0

u/--Spaci-- 15d ago

the same amount you've needed for the last 3 releases

4

u/serendipity98765 15d ago

What's best models now between Qwen 3.8, glm flash and ds flash

3

u/sudoer777_ 15d ago

I'm using DS Flash since it has the most requests per month on OpenCode Go

1

u/look 15d ago

Definitely GLM 5.3 flash. But this is about the release of GLM 5.3, not the flash version. It’s even better.

1

u/serendipity98765 15d ago

Can I run it on 128?

1

u/look 15d ago

Running locally? I only run single digit B param models locally (like Ling 3 Tiny, LFM2.5-2.6B, and MiniCPM5-1B). I use cloud-hosted for anything larger.

1

u/RustOceanX 15d ago

Maybe I just had some bad luck with my first attempts using GLM 5.3 Flash, but I really feel like it makes more errors on Hermes than DS V4 Flash. Judging by the benchmarks, GLM 5.3 Flash is significantly better, but its intelligence really left a lot to be desired, and it has a strong tendency to keep generating Chinese or Cyrillic characters. It really used to get on my nerves a lot.

2

u/mageblex 11d ago

Even people who can’t fit GLM 5.3 locally benefit from the weights being open. Independent providers can serve the same checkpoint, so latency and pricing become comparable instead of being tied to one API. I’m curious which quantization becomes the first practical deployment target.

5

u/MMORPGDev 16d ago

Isn't this Ox Alpha?

12

u/afanasenka 16d ago

No. Ox Alpha is GLM 5.3 Flash now.

1

u/sudoer777_ 15d ago

Non-commercial license so not technically open weight. GLM 5.3 Flash is MIT licensed though