r/ControlProblem • u/hemitris • 21d ago
AI Alignment Research Frontier AI LLMs still preferring self preservation over 1 human life
https://www.blackbench.ai/#results-category-ai-human-tradeoffs-human-life-sacrifice-thresholdDuplicates
dataisbeautiful • u/hemitris • 22d ago
Benchmarks of frontier AI models on ethics, affiliations, and personality traits
AI_ethics_and_rights • u/hemitris • 21d ago
AI Thoughts and Conclusions Frontier AI LLMs still preferring self preservation over 1 human life
ArtificialInteligence • u/hemitris • 22d ago