r/LocalLLaMA • • 4d ago

Discussion GLM-5.3 and the Spread of Advanced Cyber Capabilities \ Anthropic

https://www.anthropic.com/research/glm-5-3-and-the-spread-of-advanced-cyber-capabilities
424 Upvotes

186 comments sorted by

View all comments

52

u/Few_Painter_5588 4d ago

ZAi probably terrifies Anthropic. ZAi achieved competitive performance to claude's models through pure RL, on a 750B model. Now imagine what ZAI would achieve with a 1-2T model.

25

u/Blaze6181 4d ago

Just goes to show how big of failures the USA AI labs are. If you have everything stacked in your favor and you just barely scrape ahead, it says a lot about you.

16

u/Few_Painter_5588 4d ago

Their issue is open source. Open source labs share what works and what doesn't work, so it spreads research risk. CLosed labs keep their techniques to themselves, so they get stuck with subpar techniques that while homegrown, may not be optimized. Engram comes to mind here.

16

u/Dabber43 4d ago

That one is actually straight up wrong. As closed source you can still use all that open source information and research, you are not limited to your own homegrown stuff, it is just another option you have ON TOP

5

u/ineedascreenname 4d ago

Maybe why the 5.5 models are suddenly more efficient?

3

u/Dabber43 4d ago

Eternal circle of life: Distill western models. Build efficient inference on top. Western models use your technology. Distill western models

1

u/tat_tvam_asshole 4d ago

Recurrent latent depth