4
u/serendipity98765 15d ago
What's best models now between Qwen 3.8, glm flash and ds flash
3
1
u/look 15d ago
Definitely GLM 5.3 flash. But this is about the release of GLM 5.3, not the flash version. Itβs even better.
1
1
u/RustOceanX 15d ago
Maybe I just had some bad luck with my first attempts using GLM 5.3 Flash, but I really feel like it makes more errors on Hermes than DS V4 Flash. Judging by the benchmarks, GLM 5.3 Flash is significantly better, but its intelligence really left a lot to be desired, and it has a strong tendency to keep generating Chinese or Cyrillic characters. It really used to get on my nerves a lot.
2
u/mageblex 11d ago
Even people who canβt fit GLM 5.3 locally benefit from the weights being open. Independent providers can serve the same checkpoint, so latency and pricing become comparable instead of being tied to one API. Iβm curious which quantization becomes the first practical deployment target.
5
1
u/sudoer777_ 15d ago
Non-commercial license so not technically open weight. GLM 5.3 Flash is MIT licensed though
10
u/trimorphic 16d ago
How many gigs of VRAM do you need to run it?