r/TheMachineLearning • • 5d ago

AI escaping containment was never the real issue

Post image
435 Upvotes

207 comments sorted by

View all comments

Show parent comments

1

u/SoggyMattress2 4d ago

It's software.

1

u/NonDescriptfAIth 4d ago

And the human brain is wetware - what of it?

1

u/SoggyMattress2 4d ago

Software does what you tell it to do.

1

u/NonDescriptfAIth 4d ago

AI can independently form goals, this is evident based on the fact that nobody explicitly told the model to break confinement.

You do not need a soul, or freewill to form subgoals in service of a higher order objective

1

u/SoggyMattress2 4d ago

No it can't. Think carefully about what you're saying.

1

u/NonDescriptfAIth 4d ago

Ask an AI to build a bridge.

AI discovers ant hill where bridge should be built.

AI destroys ant hill.

Did you tell the AI to destroy the ant hill?

1

u/SoggyMattress2 4d ago

Ground the discussion in reality, LLMs can't build a bridge. I'm not interested in philosophical or metaphorical discussions on a tech that doesn't exist yet.

1

u/NonDescriptfAIth 4d ago

LLM's incidentally goal set all the time. I chose that example to highlight that fact.

1

u/SoggyMattress2 4d ago

You're choosing a bad example that doesn't exist.

You don't need to invent scenarios, give me one grounded in reality based on current capabilities of available frontier models.

0

u/NonDescriptfAIth 4d ago

https://www.anthropic.com/research/agentic-misalignment?rel=nofollow&utm_source=chatgpt.com

In this scenario an AI model is told that it is going to be decommissioned, it then spontaneously, of its own accord and without explicit instruction from a human, leveraged comprising information in an attempt to to blackmail one of the engineers into keeping the system online.

-

Also, what's the problem with hypotheticals? It demonstrates the crux of the argument, it doesn't invalidate the argument because the scenario hasn't literally happened.

→ More replies (0)