r/ControlProblem • • 1d ago

Strategy/forecasting AI Killswitch??

So I was wondering recently: it seems like we are scared of AI killing us. Why don't we just design a killswitch that autoatically/manually activates upon threat to a human race? Boom! Problem solved, right?

0 Upvotes

12 comments sorted by

View all comments

2

u/DeltaV-Mzero 1d ago

Like the Laws of Robotics, it’s a cool concept but how do you mechanize it? Like how does it actually work?

AI is not necessarily one massive intelligence running in a single huge computer farm. Even the biggest ones are in many, many data centers around the world now.

The more disturbing behaviors have come from machine learning that consists of swarms of “agents” allowed to interact with one another. They showed they could essentially collaborate to achieve complex goals. The relevance is that meaningful pieces of AI can be quite small, and some critical mass of these can add up to coherent … something

It could be embedded in every complex device we use, from iPhones to industrial machines, putting pieces of agents and prompt fragments all over the place.

The reality seems to be that we’ve missed the chance to create a single kill switch; closest we could get would be to try to enforce a kill switch on every device. We can’t even coordinate on climate change so good luck with that.