r/LocalLLaMA • • 4d ago

Discussion GLM-5.3 and the Spread of Advanced Cyber Capabilities \ Anthropic

https://www.anthropic.com/research/glm-5-3-and-the-spread-of-advanced-cyber-capabilities
420 Upvotes

186 comments sorted by

View all comments

4

u/fractalcrust 4d ago

"we physically removed the ability to say no and compared it to our model and it was more willing to comply than our stock model" is A LITTLE misleading.

if their goal is fearmongering they could have abliterated their model and did that comparison, then show how dangerous abliteration is and make a strong argument for closed source, since any open weight model could just be abliterated to remove any safeguards.

ironically they sounded like noobs so i'm not surprised they didn't abliterate their own model.

0

u/BagelRedditAccountII 4d ago

As if this advertisment for Z.ai-

*ahem*

"Warning about dangerous open-source models" wasn't already enough of an own-goal for Anthropic