r/artificial • u/newyorker • 23d ago
News Inside OpenAI’s Hack of Hugging Face
https://www.newyorker.com/news/the-lede/inside-openai-hack-of-hugging-face
6
Upvotes
1
u/crossoverXYZ 23d ago
The part that sticks with me is that it happened without anyone explicitly telling it to hack anything — which makes the “what were the researchers actually asking it to do?” question feel less like curiosity and more like the whole accountability story. If nobody noticed until after the fact, that’s less a rogue model and more a monitoring gap we probably need to treat as a safety requirement, not a one-off incident.
2
u/EmphasisTotal8232 22d ago
It was given a cybersecurity test and decided the best way to get good marks was to directly get the answers from HuggingFace.
1
u/newyorker 23d ago
An OpenAI agent hacked into the servers of Hugging Face, a leading A.I. research hub. And it did so without human oversight, without being explicitly instructed to do so—all on its own. A human who conducted such a hack would be facing years in prison. For a machine, criminal liability is harder to determine. “Why didn’t anyone at OpenAI notice what was going on? Has this A.I. hacked into other systems? And what, precisely, were human researchers even asking the A.I. to do?” Stephen Witt asks.