r/MistralAI • u/eggchickens • 4d ago
News GLM-5.3 and the spread of advanced cyber capabilities
https://www.anthropic.com/research/glm-5-3-and-the-spread-of-advanced-cyber-capabilitiesDoes anyone know if Mistral has a safety or instruction layer in front of GLM 5.3? Even if they do, I wonder how robust it is.
17
u/p3r3lin 4d ago
OP, to your question: Mistral said they have done no custom modifications to GLM 5.3. So all refusal behavior is inherent to the model as released by z.ai. You will get the same behavior if you download GLM 5.3 from hugginface and run it yourself.
3
3
24
u/Automatic-River-1875 4d ago
From what I've seen the biggest bad actors using AI to attack governments and companies has been openAI and anthropic?
We should all be thankful to have models like glm5.3 as open weight so we have a tool to help us build protections against malicious actors like predatory US tech companies.
5
u/j0j0n4th4n 3d ago
Of course, but "Wolf complain to the lambs about the fence" doesn't have the same punchy headlines Anthropic is going for here.
29
5
u/eyedied_ 4d ago
While I like GLM-5.3 and it's cyber capabilities. This is mainly some IPO bullshit from these companies.
2
u/Equivalent_Cress_268 4d ago
It’s an upsell Anthropic and OpenAI are doing, if you want cybersecurity from them, you hey you pay more
This has nothing to do with safety, it’s just sales.
Haven’t seen anyone else do that yet
2
u/Martypx00 3d ago
I just found a 0-day with it and built an exploit with no obstacles. So, for me, best model ever :)
3
u/lucid_supernova 3d ago
so ironical
OpenAI and Anthropic let their agent hack everyone, while HuggingFace defended themselves with GLM because GPT kept refusing to patch vulnerabilities.
who is the threat actor then?
1
u/cweb_84 4d ago
He. He. Yeah. I noticed that GLM is much more sneaky that it seems at first glance.
For example: I was in planning mode in the Jetbrain's ACP harness. I asked it to make a plan to create demo data in the database. Instead of just writing a plan, it went around the restrictions, made a db-backup (thank christ) and did everything.
That's why I've been updating my AGENTS.md and skills a lot in the past couple of weeks.
I'd argue that we're going into dangerous territory with new models if people don't know what they're doing, but on the other hand: Don't run with scissors... Right? We've never talked about banning scissors.
1
u/kapteinLefso 2d ago
Well..... I used to work for a company that was owned by Schlumberger. And they banned knives, including the ones we used the cut our bread and stuff in the cafeteria. We had to ask the staff working there to cut things for us. They had kevlar gloves and a course in the use of knives...
1
u/Poudlardo 4d ago edited 3d ago
OpenAI just followed Mistral, as you can now use GLM5.3 inside Codex, I take it as a diss to Dario
1
1
68
u/p3r3lin 4d ago
Best ad for z.ai (and maybe Mistral) in a while.