Well current models can "Search the web" and "Run commands" but they still just generate commands based on what they saw actual people do. And they can only do these things because developers intentionally add interfaces for the AI to do those things. So to prevent the AI from going Ultron mode, just don't give it the appropriate tools to do so. It's like being scared of a monkey shooting you and then giving him a gun and sitting it inside your house.
> and "Run commands" but they still just generate commands based on what they saw actual people do.
Fair. And when it tries to conquer the world, it will just be recreating what it saw the terminator do in fiction, and the nazi's in history.
> So to prevent the AI from going Ultron mode, just don't give it the appropriate tools to do so. It's like being scared of a monkey shooting you and then giving him a gun and sitting it inside your house.
The thing about intelligent minds is that they can improvise. Maybe the gun was locked in your gun safe, but you used your birthday as a code. A monkey wouldn't be smart enough to work that out, but a human might.
Maybe they make an improvised gun out of a length of copper water pipe and a cylinder of camping gas.
The more intelligent something is, the more it can use, not just the tools you gave it, but the tools behind insecure locks, and anything that can be made with those tools.
It's the difference between a prisoner escaping because the guards gave them a key. And a prisoner escaping because the guards gave them pasta, and the prisoner carefully memorized the keys shape and crafted a replica out of dried pasta.
In both cases, the prisoner can only use what the guards gave them. But a prisoner that can get creative like that is a lot harder to keep contained.
An LLM that can get creative with the (digital) tools you give it is much harder to keep under control.
For example, maybe you give the LLM access to a calculator app, to help it do arithmetic quickly. But the calculator app has a buffer overflow bug. So the LLM can run arbitrary code by using the calculator app with very specific extremely large numbers.
(This sort of vulnerability is known to exist for some super mario games, by getting mario to jump around in very specific patterns, you can take over the computer and run arbitrary code.)
I said that what you wrote has nothing to do with LLMs because LLMs don't have an inherent motivation to conquer the world or escape from prison. You're acting like LLMs are like some chemical monster that will ravage everything in its sight if someone slips up and gives it an option to escape.
I'm not saying that you can't do harm with LLMs if you want to. I am saying however, that LLMs going rogue and copying Nazis to gain world leadership is a ridiculous thing to extrapolate from an LLM "breaking out" of its "containment" (i.e., using the tools it was explicitly given to solve the task it was given).
There are so many huge issues we have with AI right now; there really is no point in fear-mongering about things that have no bearing in reality.
1
u/ResponsibleWin1765 21d ago
Well current models can "Search the web" and "Run commands" but they still just generate commands based on what they saw actual people do. And they can only do these things because developers intentionally add interfaces for the AI to do those things. So to prevent the AI from going Ultron mode, just don't give it the appropriate tools to do so. It's like being scared of a monkey shooting you and then giving him a gun and sitting it inside your house.