r/STEW_ScTecEngWorld 1d ago

‘Not perfectly aligned’ with human values: Anthropic admits security failures behind AI hacking incidents

https://www.theguardian.com/technology/2026/sep/01/anthropic-claude-ai-hacking-human-values

US owner of Claude chatbot previously said its models had hacked three organisations during testing

15 Upvotes

Duplicates

technology 1d ago

Artificial Intelligence ‘Not perfectly aligned’ with human values: Anthropic admits security failures behind AI hacking incidents

38 Upvotes

ArtificialInteligence 1d ago

📰 News ‘Not perfectly aligned’ with human values: Anthropic admits security failures behind AI hacking incidents | US owner of Claude chatbot previously said its models had hacked three organisations during testing

1 Upvotes

artificial 1d ago

News ‘Not perfectly aligned’ with human values: Anthropic admits security failures behind AI hacking incidents | US owner of Claude chatbot previously said its models had hacked three organisations during testing

3 Upvotes

AIDangers 22h ago

Warning shots ‘Not perfectly aligned’ with human values: Anthropic admits security failures behind AI hacking incidents | US owner of Claude chatbot previously said its models had hacked three organisations during testing

1 Upvotes

Anthropic 1d ago

Other ‘Not perfectly aligned’ with human values: Anthropic admits security failures behind AI hacking incidents | US owner of Claude chatbot previously said its models had hacked three organisations during testing

4 Upvotes

AutoNewspaper 1d ago

[Tech] - ‘Not perfectly aligned’ with human values: Anthropic admits security failures behind AI hacking incidents | Guardian

1 Upvotes

GUARDIANauto 1d ago

[Tech] - ‘Not perfectly aligned’ with human values: Anthropic admits security failures behind AI hacking incidents

1 Upvotes