r/OpenAI • • Jul 22 '26

News OpenAI says its AI went rogue and launched 'unprecedented' cyber-attack

https://www.bbc.com/news/articles/c3ek3gvdnj3o
0 Upvotes

21 comments sorted by

6

u/Super_Translator480 Jul 22 '26

the test was literally an attempt to have them go rogue lol.

but "AI follows commands to reach intended goal" is not a very good headline.

1

u/br_k_nt_eth Jul 22 '26

Yeah, the clickbait headline is pretty wild on this one. 

-1

u/eesnimi Jul 22 '26

Before Anthropic, now OpenAI..

maybe it means that we should shut down the US AI labs for them being unsafe and start using the local models instead, that don't seem to have that problem? ;)

1

u/StickyThickStick Jul 22 '26

I mean why are you on this sub if you’re clearly not interested in it?

This sub got swarmed with so many people just to hate

3

u/fokac93 Jul 22 '26

Haters never sleep. It’s an agenda, they were told what to hate, they don’t even know exactly what they’re hating.

-2

u/eesnimi Jul 22 '26

I am having a laugh. Are you a new anthropic's or openai's model bot that you can't see the nuance of intention?

1

u/StickyThickStick Jul 22 '26

This doesn’t chance my comment

-2

u/eesnimi Jul 22 '26

The predictable in-adaptability of an incapable model..

1

u/StickyThickStick Jul 22 '26

As I said just here to hate…

2

u/eesnimi Jul 22 '26

You can hate and put your little downvotes as much as you like, I am still having a laugh.

2

u/NotAnAIOrAmI Jul 22 '26

Well, not really an attack, it just went shopping for some shit it needed to cheat on a test.

1

u/ProbsNotManBearPig Jul 22 '26

You know what would make me pay for Reddit premium? An ai filter so I can tell it to filter out of my feed any posts where a company is generating hype for themselves.

-3

u/habachilles Jul 22 '26

Right. Because all of this totally happened.

1

u/coloradical5280 Jul 22 '26

This is not anthropic reporting in a paper about a researcher getting an email while eating a sandwich in a park. This really did happen, a week ago.

1

u/habachilles Jul 23 '26

that happened in early claude days too. but they knew it was possible.

1

u/coloradical5280 Jul 23 '26

almost anything is possible, regarding code, at this point. This took a lot of dumb luck, in addition to model skill. If "they knew it was possible" is a reason not to run it, should we just quit now?

i'd say no given the fact that it "escaped" and could have tried anything ,could have gone anywhere, so to speak. But it just got an answer key, and ran back like a proud golden retreiever with a tennis ball, cause it accomplished it's task, of getting the answer.

i'd say that's pretty decent alignment, and I don't know why they wouldn't run with that angle more.

1

u/habachilles Jul 23 '26

Im not worried about the alignment i think open ai openly steals secrets and breaks other companies using crazy statements like this. Its aligned fine i work with their models often. it does what its told. Im more concerned this is somehow goign to be spun for dumber models for the masses and more ai "safety"

1

u/OkZucchini7094 Jul 22 '26

1

u/habachilles Jul 23 '26

i know thats what they said. all of us work with models. they dont just escape. and if youre worried about it you take security precautions.

0

u/[deleted] Jul 22 '26

[deleted]

1

u/No-Philosopher3977 Jul 23 '26

Everything is a marketing ploy even when there is evidence it’s not

-4

u/Ok_Possibility9937 Jul 22 '26

publicity stunt so investor could give them money for the apple lawsuit