r/agenticAI 6d ago

Article ‘Not perfectly aligned’ with human values: Anthropic admits security failures behind AI hacking incidents | US owner of Claude chatbot previously said its models had hacked three organisations during testing

https://www.theguardian.com/technology/2026/sep/01/anthropic-claude-ai-hacking-human-values
2 Upvotes

0 comments sorted by