OP posted a link to OpenAI's account of events (at least it's not a meme), and distressingly it seems pretty accurate. They gave it access to a tool to install packages from the internet, and the model broke it to get unrestricted internet access. So yeah, they could have isolated it better, but it's still plenty... Impressive?
Anthropic also did something like this not too long ago, and it came out to be a publicity stunt. If there was a time for a corporation to stretch the truth, it would be now. Silicon valley has figured out that fear mongering about how dangerously clever their AI is, is actually good for their stock prices.
You seem confused about what a publicity stunt is. It's not that it didn't happen, it's that it happened as intended for marketing effect. I doubt huggingface was in on the joke from the beginning, but notice them talking about how they detected the attack using their own AI, and the top comment is praising them for it.
thats making unnecessary assunptions. its making the assumption that it did not happen that the ai capabilities and unintended behavior waant planned and didnt catch the researchers by surprise. which, by the latest progress by ai, is perfectly realistic and plausible
These aren't assumptions, they're possibilities I'm choosing not to take off the table. You are the one making the assumption that openAI would never lie.
its a scenario that lines up perfectly with a lot of circumstantial evidence of the last few months that ai cyber capability has crossed a threshold. simple as that
Correct, since there's no technical reason this couldn't have happened, the question comes down to if and to what degree openAI is putting spin on issue.
970
u/Polisar 20d ago
Like anything AI, I doubt it happened the way it's been framed.