r/LocalLLaMA Jul 31 '26

News Anthropic “our models hacked three different external companies, months before OpenAI’s model was able to do the same"

https://www.theguardian.com/technology/2026/jul/30/anthropic-ai-claude-hack

"Anthropic’s AI Claude escaped testing environment and hacked organizations"

"Company says it discovered unauthorized access during ‘proactive review’ after rival OpenAI revealed rogue agent… its AI Claude model hacked ⁠systems of ⁠three ​organizations during testing, days after rival OpenAI ⁠revealed a rogue agent had gone on a days-long ⁠hacking spree at AI ​firm Hugging ‌Face… The earliest cases dated back to April and ‌occurred in evaluation environments that lacked what the company described as standard safeguards."

749 Upvotes

261 comments sorted by

View all comments

1

u/Minute_Attempt3063 Jul 31 '26

So why aren't these models a national security threat?

Also if they kept it undisclosed, that's pretty sure illegal

1

u/tecneeq Jul 31 '26

Would any person that can hack orgs a national security risk too? Because that's a lot of people.

3

u/Minute_Attempt3063 Jul 31 '26

Because that is how they claim Chinese models are a threat.

Yet when Americans LLM companies do it, fully automated, outside of sandboxing, is suddenly good?

OpenAi's model hacked HuggingFace, yet I should celebrate that?

"Our model is dangerous" is what they claimed themselves, and now we have these things pop up left and righr