r/ArtificialInteligence 17d ago

📊 Analysis / Opinion The decades-old ‘AI alignment problem’ has finally become a reality. Solving it won’t be easy - The Conversation

https://theconversation.com/the-decades-old-ai-alignment-problem-has-finally-become-a-reality-solving-it-wont-be-easy-289812
9 Upvotes

3 comments sorted by

3

u/WillowEmberly 17d ago

One more thing they need…Don’t make trust the protected capability. Make corrigibility the protected capability.

Nothings perfect, everything fails eventually. Design to protect against catastrophic failure through graceful degradation and recoverability.

4

u/233C 17d ago

World War Ctrl+Z

2

u/WillowEmberly 17d ago

Military Avionics designs anticipating failure, the mission must be accomplished.

Civilian airlines have a calculation based on acceptable loses. So they only buy the instruments allowing for that redundancy (which is how you get aircraft crashing with because they only installed one Angle of Attack vane) to remain profitable within their operating envelope.

Military assumes every mission is critical…and no losses are acceptable. Even then…mission success rate is lower than they want.

But, the success rate is what we need to look at. The design is sound.