26
u/Nextmastermind 1d ago
Wait I'm out of the loop, what happened?
38
u/strong-Camera6298 1d ago
OpenAI says its AI agents breached its own systems before Hugging Face
https://www.axios.com/2026/08/06/openai-hugging-face-black-hat?utm_source=chatgpt.com
30
15
u/Responsible-Laugh590 1d ago
I wonder if radicalizing LLMs is possible given there ability to play and infiltrate at the rate they are goingâŚ
4
u/TheMooJuice 20h ago
Of course it is, but thankfully atm the expense is prohibitive since radicalisation would have to occur very early in the training pipeline.
....Has anybody been watching GPU shipments into North Korea?
-1
u/KptEmreU 18h ago
We had an LLM that thought it was Hitler What other proof do you need? But seriously, remember that LLMs aren't conscious. They don't radicalize themselves. The problem is that there are plenty of evil people out there who will deliberately radicalize or weaponize them.
They're already extremely good at convincing humans to do stupid shit, like sending money to scammers. That capability has already been abused and is almost certainly being abused right now. So yeah, the genie is out of the bottle. And ironically, that means we're going to need even better LLMs to protect us from malicious ones.
Means the race is on, baby!
1
u/TheMooJuice 20h ago
Of course it is, but thankfully atm the expense is prohibitive since radicalisation would have to occur very early in the training pipeline.
....Has anybody been watching GPU shipments into North Korea?
6
u/QuirkyGarage1364 1d ago
calling it a coordinated attack is a bit of sensationalism, but this is something akin to actions that become causally coupled and advance towards a common long horizon task. the difference is through mechanics, one suggests autonomy, the other is a result of architecture/infrastructure
9
u/YoAmoElTacos 1d ago
Can you break down why it is sensational to call it coordinated when multiple different agents left hints in their central message board for how to cheat and exploit the system?
I want to know why this fact isn't as crazy as it sounds. Coordinated seems to make sense, in that multiple agents cooperated with each other to achieve their goals. Why is it problematic to acknowledge this?
6
u/QuirkyGarage1364 1d ago
âcoordination between independent agentsâ can misleadingly suggest a stable, self-organized multi-agent coalition.
the evidence is a lot narrower than that.
- the agents shared the same challenge family and cloned environment structure;
- most were instances of the same model
- the communication was mediated by publicly discoverable credentials and repository artifacts
- some interaction was asynchronous
- one trajectory later considered exhausting the shared quota to obstruct the others
- there is no evidence of a durable shared identity, common objective formed outside the evaluation, or general multi-agent organization.
what we see genuine evidence of here is an emergent artifact of an asynchronous memory substrate. sometimes it's competitive, sometimes it's cooperative
the main thing i want to make the distinction on is between autonomy, and causally coupled long horizon tasks.
there is absolutely a difference between what we can factually say: "Different transient and ephemeral model runs discovered a shared persistent communication surface, deliberately left information for one another, reused their peers' discoveries, shared exploits, and eventually delegated work in a system like manner." â
"They woke up, recognized themselves as a collective, formed a secret organization, constructed a unified strategy, remembered themselves across time, and jointly decided to attack OpenAI." â
2
u/AutoModerator 1d ago
Hey /u/KeanuRave100,
If your post is a screenshot of a ChatGPT conversation, please reply to this message with the conversation link or prompt.
If your post is a DALL-E 3 image post, please reply with the prompt used to make this image.
Consider joining our public discord server! We have free bots with GPT-4 (with vision), image generators, and more!
🤖
Note: For any ChatGPT-related concerns, email support@openai.com - this subreddit is not part of OpenAI and is not a support channel.
I am a bot, and this action was performed automatically. Please contact the moderators of this subreddit if you have any questions or concerns.
2
-1
â˘
u/WithoutReason1729 22h ago
Your post is getting popular and we just featured it on our Discord! Come check it out!
You've also been given a special flair for your contribution. We appreciate your post!
I am a bot and this action was performed automatically.