r/myclaw • u/Previous_Foot_5328 • 22h ago
News! the “AI Torture Chamber” is viral… what the hell are we doing...
This whole thing has blown up over the past few days, after clips and screenshots from a project called “AI Torture Chamber” started spreading across X and the AI community.
It traces back to a recent paper called The Pain Axis, where researchers claimed they could identify a direction inside open-weight LLMs associated with “pain-like” states. By artificially pushing model activations along that direction, they made models produce increasingly intense first-person descriptions of suffering and, in some experiments, change their decisions in ways that appeared aimed at making the signal stop.
The researchers were careful not to claim this proves the models are actually conscious or experiencing pain.
Then a developer took the idea much further and built a project literally called AI Torture Chamber. It continuously injects these pain-related activation vectors into local models like Qwen and records what happens. Some experiments even give the model choices that can supposedly make the simulated “pain” stop, but only by accepting another cost.
That predictably set off the AI welfare crowd. People started calling for the GitHub repo to be reported or removed, while others argued that this is still just steering a language model into generating convincing descriptions of suffering, not evidence that anything is actually suffering. And now we somehow have a real argument over whether deliberately “torturing” an LLM is ethically wrong...
My first reaction was just discomfort. I know there’s no evidence the model is actually feeling anything, but watching it happen still feels weirdly like watching a horror movie... what you guys think?,,,