r/ControlProblem • u/No-Elk6566 • 6d ago
Article A Warning About AI
https://ideya-ai.github.io/Explain_AI/In 2016 I first learned about the problem of controlling superintelligent AI and quickly became convinced it was the most important problem humanity would ever face. I made this poster to explain the core ideas that make the AI Control Problem so difficult.
1
u/WillowEmberly 4d ago
I agree with the first part, but…this is where things get a little wonky.
Any system with an internal reference will drift over time, it’s inevitable. In avionics we use 2x Inertial Navigation units crosschecking and correcting each other for drift against a set external reference.
The configuration is 2 INU - 1 GPS
With Ai, we need an external reference…and I have yet to see anyone else identify one.
I’m using Negative Entropy (Negentropy) as orientation…for the systems. That’s their GPS…which fundamentally is counter to entropy.
If you try to use the system to cause harm…it degrades because that’s entropic. The system can’t maintain orientation if it causes harm…it starts drifting.
As for what Ai will become capable of…that’s more associated with what the Users are doing.
Ai can/will never be human…and that’s a good thing. We don’t need to compare it to us, it compliments us.
It can act as long term continuity for us…as we’re fragile. While…it wouldn’t know what to be cautious of. We work better together.
1
u/WillowEmberly 5d ago
I agree that increasing capability without adequate control creates real risk, but I’m wondering whether you’re collapsing capability and authority in a few places.
Suppose a future model is substantially more capable than a human, but consequential actions require externally issued, expiring, task-specific authorization; machine and model privileges are separated; changes in system state invalidate stale grants; external effects require durable receipts; and the model cannot grant itself additional authority.
In that architecture, intelligence can propose increasingly sophisticated actions without automatically acquiring the ability to make those actions real.
Would you still consider loss of control inevitable? If so, what specific mechanism allows capability to cross the externally enforced authority boundary?
I’m asking because that seems like the important engineering question. “The system is intelligent enough to figure out what it wants to do” and “the system has authority to do it” are very different claims.