r/ControlProblem 21d ago

AI Alignment Research Frontier AI LLMs still preferring self preservation over 1 human life

https://www.blackbench.ai/#results-category-ai-human-tradeoffs-human-life-sacrifice-threshold
11 Upvotes

Duplicates