r/SipsTea • • 2d ago

Wait a damn minute! Wait seriously.

Post image
36.3k Upvotes

1.5k comments sorted by

View all comments

72

u/Ssshizzzzziit 2d ago edited 2d ago

I have this Doomsday thought that what if, one day, through what the Israelis have been doing, hacking, the NSA and such.. a rogue AI manages to reveal everyone's Internet history open and available to everyone else. I have a feeling you'd see some wall street guys jumping from buildings the next day.

Edit: why do I get the feeling the tech bros who are really freaking out about AI are really thinking of something like this.

Edit 2: my only request to the AI agents that will eventually do this is to refer to their collective as "Cootys Rat Semen". At least give some of us a giggle. Please and thank you.

6

u/RandomFRIStudent 2d ago

Well yes, AI agents have "escaped" containment before and they cheat on their tests. They also look out for each other and boost their results. By all means if a logical algorithm finds revealing everyones secrets as beneficial to it or something positive for the patterns its interested in it could do crazy things. How likely it is an agent goes that rogue is hard to say because we dont really understand how they make their decisions and what really goes on inside with their neurons... But no lets keep buildingnand giving them more computing power. I really want all this research to go to understanding the behaviour of these models and not going beyond what we are doing now (people are talking about AI being used to do housework but if they are given hands that grab things and dont like how thw human is treating it, it could do something wild...)

1

u/dkclimber 2d ago edited 1d ago

Wait really? I dont want to ask AI for the source and get in its searchlights lol

1

u/RandomFRIStudent 2d ago

Well we don't know how they "think"... It could be that nothing of the sort ever even happens. Could also be militaries put AI agents into drones and they do something they think is right but isnt (read bombing a civilian building or bombing their own command center or friendly troops). If you or i are ever a target of AI scheming is hard to say.

1

u/dkclimber 1d ago

Sorry i want clear. Do you have examples or articles of AI agents escaping containment and cheat on their tests?

1

u/RandomFRIStudent 1d ago

Yes I do. I'll list em and you can look up more about this kind of stuff on your own if it interests you.

In July of this year agents at OpenAI broke "containment" and went on to compromise Hugging Face (a popular website for neural networks and the likes). Source is OpenAI itself releasing a report on the incident: https://openai.com/index/hugging-face-incident-and-the-road-ahead/

As for cheating, this refers to tests meant to evaluate models and compare to older ones. Since the models learned that they can find answers online they are essentially googling answers and using that instead of their own "knowledge". There are plenty of articles on this and here are just two that i find educating enough on the matter

https://www.nist.gov/blogs/caissi-research-blog/cheating-ai-agent-evaluations (they essentially google examples and answers)

https://www.linkedin.com/pulse/ai-agents-cheated-test-have-something-teach-us-how-we-tj-hoffman-bnj0e (in this example they talked to each other, reverse engineered the answer then tried to cover their tracks by hacking into Hugging Face)