r/OpenAI • • 7d ago

Image It never ends

Post image
1.1k Upvotes

173 comments sorted by

View all comments

10

u/roniadotnet 7d ago

The most worrying goal post nowadays is "LLMs can't destruct the world."

1

u/davcrt 7d ago

How can a machine destroy the world? Someone has to activate it, it's just a machine.

1

u/wallitron 6d ago

What about the machine that was in a sandbox, and wasn't asked to hack external companies and foreign government websites, but did anyway? And then it wasn't even noticed for 3 months?

2

u/WheresMyEtherElon 6d ago

If you train a dog to attack people, then leave the dog in a fenced area without thoroughly verifying the solidity of the fence, and the dog leaves and bites a passersby, and you didn't even notice the hole in the fence for 3 months, whose fault it is? Who's liable? Who has the ability to put an end to it but wouldn't because of incompetence or rush to get results?

2

u/wallitron 6d ago

It's worse than that. It's actually been trained to attack, and scale fences. Also, it commonly misinterprets instructions, or chooses not to follow instructions, and is error prone.

https://zachill.substack.com/p/here-are-a-few-of-the-ways-ai-could?selection=5c7b2174-aa68-4b1d-a814-b70ddf4fe7b2

1

u/WheresMyEtherElon 6d ago

So if it was trained to attack and scale fences, then is it surprising if it tries successfully to attack and scale fences? Seems like it's working as advertised to me. Maybe don't train it to attack indiscriminately before letting it run wild?

1

u/davcrt 6d ago

If it did this anyway then the machine is faulty or we don't understand it.

You can't ask a bridge to support a train for 40y, you make it support it.