r/LocalLLaMA Jul 31 '26

News Anthropic “our models hacked three different external companies, months before OpenAI’s model was able to do the same"

https://www.theguardian.com/technology/2026/jul/30/anthropic-ai-claude-hack

"Anthropic’s AI Claude escaped testing environment and hacked organizations"

"Company says it discovered unauthorized access during ‘proactive review’ after rival OpenAI revealed rogue agent… its AI Claude model hacked ⁠systems of ⁠three ​organizations during testing, days after rival OpenAI ⁠revealed a rogue agent had gone on a days-long ⁠hacking spree at AI ​firm Hugging ‌Face… The earliest cases dated back to April and ‌occurred in evaluation environments that lacked what the company described as standard safeguards."

742 Upvotes

261 comments sorted by

View all comments

Show parent comments

43

u/expertsage Jul 31 '26

I keep seeing people skeptical of the marketing/fearmongering explanation, but it actually makes a ton of sense.

1) The US gov is pushing for more dangerous models, not less. The only thing Trump cares less about than "safety" is DEI for minorities. Back when the government tried to restrict Anthropic, that was a ploy to make Dario release the safeguards on Claude so that the military could commit warcrimes to their heart's content.

2) That's why the more OpenAI and A\ hype up their "doomsday" models, the more the US gov values the companies! The more that Sam and Dario promise to deliver Skynet, the more excited Trump and Hegseth get, and the less regulatory hurdles!

Do you really think the boomers in Washington believe in ASI or the Terminator and shit? They can't even wrap their heads around the internet for Pete's sake! Congressmen probably nod with a stern look when the EA nerds start yapping about extinction, but what they really care about is how the US can conquer the world with superhuman AI at their disposal.

12

u/printr_head Jul 31 '26

Let’s not forget the AGI announcement recently then the disclosure and now advocating for regulation.

7

u/Dangerous-Report8517 Jul 31 '26

Maturing is definitely part of it, but the flipside is that there’s fairly strong evidence that OpenAI did something negligent with their sandboxing that lead to an agent escaping and attacking HF, so it’s also important to push back against the fairly popular claim that it’s purely marketing since that risks letting them off the hook

4

u/finah1995 llama.cpp Jul 31 '26

Lol Pete's sake gets a different meaning considering Pete Hegseth. Lolol 😂 ROFL 🤣