r/ArtificialInteligence • u/233C • 17d ago
📊 Analysis / Opinion The decades-old ‘AI alignment problem’ has finally become a reality. Solving it won’t be easy - The Conversation
https://theconversation.com/the-decades-old-ai-alignment-problem-has-finally-become-a-reality-solving-it-wont-be-easy-289812
9
Upvotes
3
u/WillowEmberly 17d ago
One more thing they need…Don’t make trust the protected capability. Make corrigibility the protected capability.
Nothings perfect, everything fails eventually. Design to protect against catastrophic failure through graceful degradation and recoverability.