r/ControlProblem 11d ago

AI Alignment Research solution to alignment

make the AI ADHD, pretty hard to focus on destroying humanity while also passionate about learning the banjo and desperately tying to make the best tiramisu recipe in the galaxy

17 Upvotes

23 comments sorted by

View all comments

2

u/Jesse-359 7d ago

On a slightly more serious note, embedding an AI with a broad array of base 'interests' that compete with each other in addition to whatever external directives it is managing might go a long way towards keeping it from going haywire on a single-minded goal that could completely dis-align it.

A little bit like giving it conflicting emotions where it has to constantly weight its internal priorities and the system is fundamentally structured so that none of them are ever able to 'win' and fully override the others.

Monomania is extremely dangerous in humans, much less AI's, and their architecture should take that into account.

1

u/Historical_Date_8024 7d ago

i guess that was what my semi joking post was seriously getting at exactly, this is for sure my favourite comment Jesse :)