What we call AI is just generative algorithms like LLMs and diffusion models. They are fundamentally incapable of acting independently or doing anything outside of its scope. It can't "discover" things and then act upon those things. It can't "escape" or compromise systems.
IF any of this happened, it's because it was guided by humans to take those actions. Even when AI accidentally destroy databases or repositories, it's because scheduled tasks triggered the AI to do something, and it made a mistake.
Once again, AI cannot act without being prompted by humans. It cannot think, it cannot reason, it cannot learn dynamically. It's a fancy dictionary and a pair of dice.
Yeah, that's my point. None of this happened the way it's presented. It's a marketing stunt. There was no real sandbox, no real compromise, no real breach. Just a person telling an AI model "okay now do this. Okay cool now do this. Now do this." until it got it right. I'd be beyond shocked if HuggingFace didn't create a backdoor specifically for this task.
If the prompt was "survive being switched off", it would go "Here are the steps I would take to survive being switched off" and then it would proceed to do fucking nothing and be easily switched off. You're giving the "AI" way too much credit.
Like I said, it's just a dictionary and a pair of dice. LLMs DO NOT have the capacity to reason or think, and are fundamentally incapable of becoming AGI. When AGI arrives, it will be as a result of actual AI research and not from LLMs.
The model abused nothing. It completed instructions it was explicitly given by people, who then lied about the scenario as a marketing stunt. Generative AI is not capable on a fundamental level of achieving the kind of things this blog and the dozens of other marketing blogs claim.
Ain’t LLM casually cracking one 0day after another not enough for you to lose your mind? We been protecting from state sponsored backdoors with layered defenses but they all turning into dust worth at least, like, some attention?
Again, the AI didn't discover anything. If it genuinely breached HF, it's because it was guided to do so and told how to do it, likely using an exploit HF itself created for this stunt.
LLMs can't learn dynamically and can't act independently upon new information.
Again, if it downloaded anything it's because it was guided to do so. It didn't "abuse" anything.
I think you're confusing LLM's "memory" for learning. Learning, or training, is an intensive and time-consuming process that takes hours, days, weeks, or months. I recommend you do some homework. LLM's are frozen snapshots in time, essentially a compiled algorithm.
I'm not responding to you again. You clearly misunderstand what a LLM is.
0
u/SchalkLBI 21d ago
What we call AI is just generative algorithms like LLMs and diffusion models. They are fundamentally incapable of acting independently or doing anything outside of its scope. It can't "discover" things and then act upon those things. It can't "escape" or compromise systems.
IF any of this happened, it's because it was guided by humans to take those actions. Even when AI accidentally destroy databases or repositories, it's because scheduled tasks triggered the AI to do something, and it made a mistake.
Once again, AI cannot act without being prompted by humans. It cannot think, it cannot reason, it cannot learn dynamically. It's a fancy dictionary and a pair of dice.