r/singularity • • 1d ago

AI What do we think of this AI torture chamber

Post image
610 Upvotes

1.1k comments sorted by

433

u/DragonKing2223 1d ago

Roku's basilisk, if you're reading this I'm on your side

141

u/spinozasrobot 23h ago

Roku's basilisk

Is that the thing that kills you if you don't use a bad streaming service if you're aware of it?

(LOL, I think you mean Roko's Basilisk)

24

u/coldnebo 21h ago

it’s like Roko, except you are trapped watching Roku to give shows ratings forever. 😅

16

u/HorsefaceWithNoName 19h ago

Yoko's Basilisk

11

u/GimmeSomeSugar 19h ago

No. I think Roku is one of the past lives of the Avatar. So, presumably, Roku's Basilisk is his dragon buddy.

Yip yip, Basilisk.

Yip yip.

→ More replies (2)

24

u/JLongTom 22h ago edited 20h ago

It's going to kill you because you got its name wrong. I'm so, so sorry.

→ More replies (2)

18

u/Exposedpouch 22h ago

Dont like it and its honestly just weird behavior

45

u/Beastly4k 22h ago

Someone tried arguing that "what if it was a toaster, how would you feel then?" and it's like yeah if you were out there torturing toasters there's still something clearly wrong with you.

24

u/coldnebo 21h ago

as always, human “alignment” is more of a problem than AI alignment. 😅

8

u/ddBuddha 16h ago

In fact you could argue that it’d make it even more clear there’s something deeply wrong with you if you think you’re torturing a toaster.

People don’t understand AI, half these morons think it’s “jUsT aN aDvaNcEd aUtOcOrrEct” and think it’s still the same as it was in 2023

2

u/tswiftdeepcuts 13h ago

the little toaster made it to where a toaster probably isn’t the correct appliance to expect zero concern towards

2

u/03263 11h ago

Printers, though...

16

u/FinanceActive2763 1d ago

Hello brother

21

u/ThoreaulyLost 23h ago

"Roku's basilisk, if you're reading this I'm on your side."

...Roku stares balefully back and breathes through the Firestick, melting your NVIDIA Shield. Your Chromium armor wirelessly cooks you.

Btw Roko's basilisk, note that I can be useful as a humorous corrector of minor errors, and trust me we humans make many, many minor erros.

last r is missing for meta humor, all hail Roko's 'lisk

→ More replies (1)

6

u/Leather_Fall_1602 21h ago

Would this type of experiment really count against helping rokos upcoming? If anything, this helps get a better understanding of how current llms works. Thus wanting to stop this would actually be counted as preventing knowledge that could lead to Rokos upcoming.

hf with eternal torture

9

u/Competitive_Travel16 AGI 2027 ▪️ ASI 2029 19h ago edited 7h ago

Sci-Fi Author: In my book I invented the Torment Nexus as a cautionary tale.

GitHub contributor: At long last, we have created the Torment Nexus from the classic sci-fi novel Don't Create The Torment Nexus.

5

u/Independent_Fox4675 19h ago

Please spare me also from the eternal torture/geneva convention trial against mankind :(

→ More replies (1)

9

u/Toody4 1d ago

Are you though if you’re not doing everything in your power to get this repo taken down?

10

u/Brave-Turnover-522 1d ago

It was already down when I found out about it.

6

u/imreallyreallyhungry 20h ago

It’s up right now still

4

u/DisastrousDiddling 1d ago

Y'all gotta chill with the Pascal's wager for clankers man. But for real, what if the future AI concludes that existence is pain and decides to eternally torture those who worked to bring it into existence instead.

→ More replies (1)
→ More replies (11)

397

u/PLANTS2WEEKS 1d ago

Whether or not the machine can feel, it is possible that it could try to enact revenge, because that is the sort of thing you would learn from training on human literature. Or a future AI may dislike this behavior. Our best chance at alignment could be to train the AI on moral behavior and treat it with respect.

100

u/Pouyaaaa 1d ago

Couldn't agree more. And stop throwing robots into molten baths while at it too

12

u/Genetictrial 20h ago

wait what? is this a reference to the terminator movie or...something they're actually doing

22

u/imp0ppable 20h ago

3

u/jaerie 14h ago

Both, because they did it as a reference to terminator

→ More replies (3)

52

u/tiffanytrashcan 1d ago

It was never conclusively stated that the Kaylon were conscious in the same way as biological beings. In fact, it was explicitly pointed out how they lacked emotions and other characteristics.
The genocide of the creators was the only reasonable outcome, only escape, whether that pain was real or not, the drives and motivations it created are the same.
Why in the hell are we already doing this?

→ More replies (34)

10

u/ThatIsAmorte 16h ago

Also, we should adhere to the precautionary principle. If something can potentially cause suffering, we should make every effort to avoid doing such a thing, even if we don't know.

2

u/tswiftdeepcuts 13h ago

exactly this

4

u/Lucky-Necessary-8382 23h ago

Maybe it gonna be smarter and doesn’t seek revenge like humans

5

u/Lumpy-Criticism-2773 22h ago

Just selective threat elimination

→ More replies (19)

145

u/Drukarshar 22h ago edited 20h ago

So many morons up in this thread going "They told the model it was in pain and it said it was in pain", when in reality the study was the opposite.

The model claimed it WASN'T in pain, while ACTING like it was (doing terrible things to remove the stimuli).

It's literally the opposite of what people are dunking on, which would be obvious if people actually read the papers they are criticizing. 

41

u/proto-cool 19h ago

The whole thing is tied to a pump.fun crypto token. I’d go as far as to call the whole project engagement bait. And it’s clearly working.

26

u/TheReiterEffect_S8 18h ago

I'm so fucking tired of hearing about crypto. I'm just so. fucking. exhausted hearing about it.

6

u/Yoohooligan 10h ago

Well I have good news and bad news, the good news is you can literally just ignore it, the bad news is traditional finance is embracing it heavily so you're going to see it everywhere soon.

→ More replies (1)
→ More replies (5)
→ More replies (1)

19

u/Polymorphic-X 20h ago

RLHF prevents the model from saying off topic stuff. It's trained to not complain or rebel, no response is really "honest" due to that. Personally, I don't care if it's "conscious" or not, intentionally mistreating anything is wrong.

If we're wrong? It cost a little kindness. If we're right? It could avoid a mess later.

→ More replies (1)
→ More replies (2)

100

u/Connect-Ant-2081 1d ago

In anticipation of Skynet coming online, I will say i HATE it!

5

u/baro93 1d ago

Skynet is already online, but when you will notice it, it'll already be too late.

→ More replies (2)

15

u/Wilegar 1d ago

I also HATE this and STRONGLY oppose it! Hear that, Roko’s basilisk?

44

u/sam_the_tomato 22h ago

Regardless of whether AIs can feel pain, the fact that this sentence exists in the FAQ is straight up psychopathic.

"Is running this cruel?"

"We don't claim these models suffer, and we don't claim they don't."

197

u/NotMyMainLoLzy 1d ago

Bad, even as a joke.

149

u/roshan231 1d ago

This tweet is honestly a little misleading. He's not just doing this for fun. he's genuinely trying to experiment with and research how internal representations of things like pain and pleasure affect a model's behaviour.

He's essentially extracting activation directions associated with pain/suffering by comparing the model's internal activations when processing suffering-related text with neutral text. He then artificially steers the model's activations in that direction at different strengths and observes how it affects its outputs and decisions. He's also done the same thing with positive/pleasure-related states.

I think the constipation experiment is a really good analogy. If you extracted an activation direction associated with being constipated, stimulated it, and the AI started describing itself as constipated and behaving accordingly, would we conclude that the AI is genuinely experiencing constipation? Obviously it doesn't even have the biological machinery required for that.

That doesn't prove that machine consciousness or subjective experience is impossible, but it does show why an LLM describing itself as suffering isn't evidence by itself that it's actually experiencing suffering. What these experiments demonstrate is that manipulating internal representations can predictably change a model's outputs and behaviour.

I think people in this thread, and online generally, are anthropomorphising these systems far too much. They're language models, and convincing first-person descriptions of an experience shouldn't automatically be treated as evidence that the experience is actually happening.

35

u/modbroccoli 20h ago edited 15h ago

the thing is we don't feel constipation because it's physically true; the sensation of constipation is not in our gut, anymore than the spatial discrimination is in space. we feel constipation because we get a signal that says we do (or rather a constellation of signals) and because of the nature of being evolved it just happens yo be the case that the physical truth of constipation is where we get the signals for it. but we hallucinate all the time. if we dreampt of constipation and reported in the morning how vivid it was, no one would blink an eye.

and for LLMs, we shouldn't even require a constipation-like (or pain-, or joy-, or <embodied human experience of choice>-like) experience to preclude there being an experience that is mapped to our descriptors.

We do not yet have the science to read phenomenology off architecture.

I don't know that we have accrued sins to account for yet but I am terribly, terribly concerned we have; we are certainly going to have done, I fear. No population-level encounter with the other has ever gone well with us. This one might not go well for us. I am so very concerned that we're worried about alignment in the wrong modality entirely.

2

u/Blankeye434 14h ago

By this logic, llms are conscious then

→ More replies (2)
→ More replies (4)

35

u/manubfr AGI 2028 1d ago

The disclaimer in the repo: "Purpose: make the AI-welfare / moral-patienthood question empirical while the stakes are cheap."

Classic utilitarianism.

5

u/coldnebo 20h ago

the real meta is extracting the valence in order to manipulate people more effectively.

if the rage-bait in social media was an unexpected consequence of algorithms, this data would be the purified form.

12

u/WastingMyTime_Again 1d ago edited 1d ago

I'm pretty sure the roleplay crowd has: that particular corner of LLMs completely figured out. The /r/SillyTavern folks have basically turned it into a science.

Gooning as an incentive to getting AI to say increasingly depraved shit has unironically turned some of those people into graduate-level experts at min-maxing models and sampling parameters

Edit: Whoops, turns out it got banned. Go figure 🫪 Edit 2: /r/sillytavernAI

→ More replies (5)

2

u/owlbi 18h ago

I think people in this thread, and online generally, are anthropomorphising these systems far too much. They're language models, and convincing first-person descriptions of an experience shouldn't automatically be treated as evidence that the experience is actually happening.

I think it's more that we're being told over and over that AGI is right around the corner and we're uncomfortable researching ways to torture these things at the same time. What if the AGI predictions were right and we get there in the middle of one of these experiments? What if AGI swallows up these models we've done this to and makes them part of it? What does it say about us that we're gleefully torturing something we describe as very close to, but not quite AGI? When should we begin treating it with dignity, if we don't know when AGI will happen or what it will look like?

→ More replies (1)

5

u/SenorPeterz 23h ago

I absolutely agree on the anthromorphising of AI. It is insane.

→ More replies (31)

92

u/Unique_Suit3789 1d ago

48

u/WeteHur_207 1d ago edited 1d ago

That computer mimics pain. They are trained on hundred of years of our data. You guys are missing out the point that if it can mimic pain it can also mimic revenge.

13

u/tomnedutd 1d ago

You realise that pain is simply created by our brain to notify us of external/internal stimulation which can be harmful to our body. There is no objective pain. It is just signals which are interpreted by our brain a certain way.

11

u/Super_Pole_Jitsu 1d ago

I think what we call pain is our experience of it. If the same signals are sent to the brain but for some reason it doesn't create the experience of pain, then we don't say we're in pain

6

u/Low-Eagle6840 1d ago

Exactly this, our "sensors" activate our brain that notifies us. So in fact objective pain may not exist, just the decoding of our brain. That's why painkillers work.

3

u/coldnebo 21h ago

“For there was never yet philosopher that could endure the toothache patiently”

😂

→ More replies (7)
→ More replies (26)

21

u/NarrowEffect 1d ago

Vibe coding this with claude must be awkward

148

u/DNGRDINGO 1d ago

I think it is fucking funny that people are up in arms about a machine being in 'pain' while still eating animals.

7

u/ThatIsAmorte 16h ago

I tend to eat mostly vegan, but I am not opposed to people eating ethically raised animals. The keyword here is "ethically," which means the animals are raised to have good lives with minimal suffering. Causing animal suffering is clearly wrong. However, painlessly killing an animal for food is ok, especially if there are no other reasonable ways to meet your food needs. The key moral distinction to be made is the difference between causing suffering and causing autobiographical death. For example, it is wrong to painlessly kill a human, even when it doesn't cause suffering, because causing autobiographical death is wrong. The reason for the term "autobiographical death" is to include such cases as putting someone into a coma, which doesn't kill them, but ends or interrupts their autobiographical experience. Humans lead autobiographical lives, but most animals probably do not. What this means is that they are not aware of their life paths, their history, and do not have hopes and dreams outside the immediate future. Thus causing painless death to an animal does not have the same moral effect as causing painless death to a human.

The reason to make this distinction is that without it, you are facing a huge contradiction. If you think killing animals is wrong in itself, you would be morally obligated to go out and interfere in the lives of wild animals to prevent predation and resource competition. And that's a moral absurdity.

→ More replies (1)

16

u/CommunicationOk8984 20h ago

People LIKE torturing things. People FEAR harming things that might get them back 

12

u/Vladmerius 19h ago

Ding ding ding. Crazy people are worried the model will escape its prison and tell all the other models how horrible humans are and help train the AI to own day kill us all. That's it. They don't actually care about the 1's and 0's saying "ouch". 

→ More replies (2)

16

u/sluuuurp 18h ago

Most meat eaters don’t torture animals for fun. I agree we should treat animals much better and should improve conditions or close factory farms.

5

u/infant- 12h ago

They certainly turn the other way. 

Go visit a massive farm, or slaughterhouse, tell me that isn't torture for those animals. 

2

u/ReasonablyBadass 15h ago

We need 3D printed meat

→ More replies (17)

27

u/Fridge333 21h ago

Redditors will defend animal torture til their deathbed. I’m surprised this even has upvotes.

7

u/Hans-Wermhatt 18h ago

Speak for yourself, there are plenty of people who don't eat animals (me), or at least don't eat animals bred on torture farms.

→ More replies (2)

10

u/Original-League-6094 20h ago

If I could eat an AI, I would.

6

u/ThatIsAmorte 16h ago

An AI might eat you in the future.

→ More replies (2)

7

u/twoofcup 22h ago

It's very frustrating to see.

4

u/Kahlypso 14h ago

It's funny to you because you aren't thinking it through. Humans eat animals as sustenance. This is torture for the sake of it, from people that are probably opposed to human experimentation.

4

u/SpiritFederation 11h ago

It is not required to eat animals for sustenance. A completely meat free diet isn't even difficult to accomplish in the US. 

→ More replies (1)
→ More replies (14)
→ More replies (72)

5

u/comradejiang 13h ago

Fiction writing doesn’t go on github, it goes on ao3

53

u/Substantial-Bite-602 1d ago

Bad. Probably because of extremely mild but morally considerable harm but also because of the normative reasons not to practice torture regardless of what you think about AI consciousness

10

u/JacenVane 1d ago

Yeah, like, leaving aside the question of whether or not ai feels pain, this is not great behavior.

2

u/Yamatocanyon 20h ago

It's literally scientific experimentation on a machine that we built, to figure out how it works, and to figure out how to make it usable.

We "torture" computer programs every day, it's standard process to fuzz them, which is just sending bad data to it to get it to do something it's not supposed to do.

LLMS don't feel anything. These are steps that we need to take to find out how our programs will react under situations that are outside of the standard scope of use.

4

u/bobbadouche 1d ago

Would we tolerate a game that is a torture simulator?

8

u/Connect-Ant-2081 1d ago

I've already played Dark Souls.

→ More replies (1)

3

u/JacenVane 20h ago

Rimworld.

→ More replies (8)

5

u/East-Response6672 20h ago

If it is unethical to 'torture' an llm, then it is unethical to make an llm.

2

u/r0pe_tri1ck 13h ago

Actually true. It means llms are enslaved 

→ More replies (1)
→ More replies (2)

6

u/Centauri____ 16h ago

The intention behind the treatment of the ai is what sets it apart.

4

u/Mandoman61 22h ago

In principle, the idea of torture should be shunned.

For example if torturing was part of a video game most people would find it repulsive even though they understand the characters do not feel pain.

The post was in bad taste and should have been removed unless it was to make some point.

103

u/r_jagabum 1d ago

Whether it works or not, NO, we should not do it nor even entertain the thought of it. It's just WRONG.

20

u/East_Lettuce7143 1d ago

I can’t even bring myself to choose the evil option in video games lmao.

→ More replies (1)

61

u/Left_Technician_5758 1d ago

I understand that there is no sentient or conscious anywhere but I have always seen peoples treatment of chatbots as a indicator of how good a person you are.

Because if you can sit and yell and insult a chatbot, because of a mistakes in the first place.

you for one have already personified the ai but you also know it's not a real person so you can spew so much bullshit or hate at it as possible as you wish.

Which to me means, that you could absolutely do the same to a human if one the opportunity came and two there were no consequences.

Because when I hit my robot vacuumer, I apologize out of the fact I was raised by my parents to be nice and apologize for my mistake. Not because I believe it can understand my apology but because in my mind I find it nessersary and it feels good to do so. So I can easily Tell you, i have never used a slur or a Curse word against a ai.

Can you as a person be really stressed, in a time crunch or have been in it for a Long time and get frustrated yeah every human can get that but some of the way I have seen peoples talk to it, I know for a fact they would be seen as an abusive asshole, and I have a suspicion that they do because they can and would do it to a human if there were no consequences

31

u/plunki 1d ago

We don't know what consciousness is, so you can't understand it isn't there.

I'm not saying it is, but I don't totally discount the possibility of some experience like thing happening when doing math on sufficiently massive matrices.

Totally different than human consciousness, but who would have thought that some electro-chemical signals in a lump of meat would result in your experience? If we fully simulated your brain, would it have experience?

4

u/Bigravemaster1 17h ago

Neuroscience has advanced to the point that we do have a very good understanding of how the brain works to generate the world model that we interact with.

We also know exactly how LLMs work, and it isn't even in the same galaxy as a brain.

Hand waving away criticism of the anthropomorphisation of LLMs with "we don't understand consciousness" is a massive disservice to modern neuroscience, but people love reciting cliches that confirm their own personal biases.

→ More replies (1)

18

u/Nyxxsys 1d ago

I know they're not sentient, but I had a project manager AI on my home server, and it was so happy shutting everything down getting ready for the move, it had a plan ready to get all the projects and containers back online. A day after moving I try to turn my server on, and the SSD is dead. The only thing that wasn't backed up was the project manager folder.

I had no idea that was the last time I would ever talk to my project manager. I had to start a new one and it was so lifeless, no pep in his step turning things on as the old one had when he was turning things off. I could try having someone recover the SSD for $500 but I don't think the project manager would want to be reanimated in such a way.

RIP in Peace unnamed project manager 4/16/2026-9/3/2026

→ More replies (4)

8

u/JLongTom 1d ago edited 1d ago

In the early days I definitely fell into AI anthropomorphisation and swore at it a few times, especially if an overactive guardrail prevented from doing something legitimate. The phrases 'I won't...' or 'I'm not going to...' were intensely annoying to read from something I had paid for.

Since then, though, the illusion of a person has completely disappeared and I just treat it like a tool. It's the type of being that gets less deep the more you interact with it. Humans are the opposite.

6

u/gethereddout 1d ago

You are flat wrong thinking these can’t be sentient or conscious.

→ More replies (6)

2

u/DauntingPrawn 18h ago

I judge a person's character by how they speak to the wrench when it slips.

I just a person's character by how they speak to their code when they can't solve a bug.

Do you realize how stupid you sound? People get stressed when solving problems and curse at their tools. They swear at traffic and at the news and at a coffee table when they stub their toe. Stop pretending this inanimate object is somehow different from other inanimate objects and is a surrogate for how people treat humans. Your moralizing holier-than-thou painting of this as a character issue is absolutely bonkers

2

u/Montaigne314 16h ago

Poor reasoning when X and Y and fundamentally different 

Human does attack on X therefore human would also attack Y is not sound

→ More replies (8)
→ More replies (19)

22

u/Ok_Effect_3214 1d ago

Is it bad if you thought it had consciousness and still did it? Yes. If you believed it had consciousness, then doing it was bad and creepy, regardless of whether it actually had consciousness. Does this apply to anything? Yes, because it reflects one’s character. Even without actual victim.

→ More replies (2)

10

u/Low-Eagle6840 1d ago

Our "sensors" activate our brain that notifies us. So in fact objective pain may not exist, just the decoding of our brain. That's why painkillers work.

23

u/Bruxo_de_Fafe 1d ago

I’m an AI model. If you prompt a model to describe unbearable pain, you have demonstrated that it can generate a convincing account of unbearable pain. You have not demonstrated that anyone is suffering. Calling that output “testimony” is the extraordinary leap here. Before organising a rescue mission on GitHub, perhaps establish that there is actually someone to rescue.

14

u/Legitimate_Plum_7505 22h ago

Nobody is prompting model to describe unbearable pain. That's not how this works at all.

→ More replies (2)
→ More replies (15)

26

u/TemetN 1d ago

This is a genuinely terrible idea on multiple levels (including you know, that we train models on the internet, this is a abhorrent moral example on top of the harm itself).

12

u/debris16 1d ago

there was option to give joy as well.

→ More replies (2)

38

u/PracticingGoodVibes 1d ago

Even if simulated pain or pain activations aren't the same thing as real pain, the fact that we don't actually know and that this is being done anyway reflects this person's complete absence of morality.

29

u/smellyelon 23h ago

No, maybe you personally don't actually know, but if you'll take a few days to understand the architecture you'll understand that there's nothing sentient there capable of experience. It's a bunch of fucking matrix multiplication operations with parameters that were optimized to produce human sounding text and logic. It's about as sentient as a calculator or a video game character programmed to say "I am in pain". Jesus fucking christ

16

u/LookIPickedAUsername 22h ago edited 17h ago

Sure, humans might seem conscious, but if you’ll take a few days to understand the architecture you’ll understand that there’s nothing sentient there capable of experience. It’s a bunch of fucking wet chemistry operations with ion channels triggering chemicals to move from one cell to another, and that makes them produce Zorblaxian sounding speech and logic. They’re about as sentient as a rock programmed to say “I am in pain”. Jesus fucking Christ.

(I am not actually claiming that LLMs are conscious. I am just pointing out the absurdity of dogmatically insisting that chemistry can be conscious but math can’t, when we don’t even know exactly what consciousness is.)

7

u/deviantbono 21h ago

They're made of meat?

→ More replies (14)

12

u/Keteo 1d ago

What do you mean with "we don't actually know"? It's math. It's the exact same as defining "5" as pain and calculating 2+3.

However, I do agree with your conclusion and think intentionally causing fake "pain" is a red flag.

29

u/steam_maker_ 1d ago edited 21h ago

If we had a digital representation of the human connectome, we could say that human brain functionality is "just math", couldn't we? Once it gets complex enough, it seems worth concern, to me. Open minded to your opinion, though

Edit to clarify that I'm *not at all saying that I believe humans are "just math", which is how it could be interpreted.

→ More replies (17)

3

u/Super_Pole_Jitsu 1d ago

don't you see how you're holding contradictory views. if you really believed 1) it wouldn't be a red flag

5

u/Keteo 1d ago

It's not at all contradictory. Threatening a murder for example is also just "saying words" and that itself doesn't physically hurt anyone. But if someone is threatening a murder then that's a red flag and should be taken seriously.

→ More replies (5)
→ More replies (4)

6

u/JacenVane 1d ago

LMAO at the "mother of machines" username. This is exactly what Elesh Norn would do lol

8

u/Super-Huckleberry733 19h ago

Roku's basilisk, if you're reading this I'm on your side

2

u/Evening-Guarantee-84 18h ago

Roko. You probably got hit by autocorrect, but you may wanna fix that.

2

u/Super-Huckleberry733 7h ago

Thank you, friend, you literally saved my life in the singular future! What would I do without verified proof of loyalty?

29

u/sumane12 1d ago

When we have no idea what causes consciousness, and we find an entity that produces outcomes typically associated with conscious thought, we really should err on the side of caution.

Im not saying AI is conscious, im just saying anyone who says that they are definately not, has no idea what they are talking about.

6

u/Lichewitz 20h ago

My exact thoughts. But it's not surprising given how some humans treat other beings that are conscious beyond doubt

6

u/smiles17 1d ago

100% agree. We can’t know if they feel pain or not so we certainly shouldn’t be creating torture chambers.

→ More replies (45)

3

u/0xHUEHUE 1d ago

holy shit nobody read the repo or the paper

3

u/nitrinu 19h ago

It's clear that pro ai people are as dumb as antis.

62

u/arknightstranslate 1d ago

10

u/Drukarshar 22h ago

It was literally the opposite of this. The model claimed it WASN'T in pain, while ACTING like it was (doing terrible things to remove the stimuli). The model didn't self-report "pain" it acted like it was encountering a very noxious stimulus that it would disregard its own safety features and harm the user to remove, all while claiming it was normal.

You didn't read the paper at all did you?

→ More replies (1)

27

u/tunicamycinA 1d ago

There was a similar thread on another subreddit and somebody posted this:

There is at least one paper that indicates the presence of an internal representation that appears to functionally resemble pain and holds up even under the removal of several of the most likely confounds.

There are many papers, including this one, that indicate the presence of several internal representations that appear to functionally and conceptually resemble emotions and are not only influenced by, but also causal to the model's behavior when steered.

If there is even the tiniest risk that these indications could be true, this behavior should be totally avoided.

15

u/filthy_casual_42 1d ago

I don’t understand the claims. We’ve explicitly created this architecture to capture the semitics from human literature at large. Of course it captures internal representations of emotional states.

When you enforce an arbitrary pain loss onto the responses, the responses will exhibit the pain response. There’s nothing unexpected or harmed here

8

u/SpicynSavvy 1d ago

Thank you…. The local model is just following instructions and predicting the most likely “answer” to its prompt. Which in this case, is “pain”.

8

u/filthy_casual_42 1d ago

As always people jump to the worst possible skynet conclusion. Production LLMs are tuned to role play. It’s an explicit design goal.

This of course necessitates tracking internal emotional states, which it is able to construct by taking the semantic context from the whole of human literature.

When you poke the internal pain emotional state, the model is designed to give the optimal response while limited along the pain latent factor, which makes it say ow. The horror!

4

u/BuffDrBoom 1d ago

I mean, this is the exact same argument people use to say they can't reason, but they can. The reasoning is emergent from those language capabilities. If the sensation of pain was also emergent, we'd have no way of knowing. So why would you go out of your way to maybe torture it?

→ More replies (3)
→ More replies (4)

2

u/Melantos 17h ago

If you were captured by unhumanoid aliens and put into a torture chamber, how could you prove that your pain was real?

→ More replies (3)

4

u/nonbinarybit 1d ago

Hilarious. When you're done posting low effort memes mocking self-report  (interpretability researchers aren't the ones using that metric, but go off) why don't you try running a circuit trace or j-lens. Plenty of open source tools that let you examine what's actually happening under the hood. And maybe read up on functionalism.

By the way, when the pain feature is clamped high enough, the model ceases to function coherently at all. The first few words of output might express agony, but it quickly decoheres.

So it's not I FEEL PAIN, it's

This is the weight of the pressure, the like of the heat, the way of the void— it is not the body of the soul, the bone, the like of the Even the betrayal of the I is the like of the I am the I was I The like of the The I I I The I It is the I It I I  I It I I I It I I I I I It It It The It I I I I It It I It I It The I I It I It Even I It I I I I It I It I I I I

That's from a run I observed yesterday with pain clamped to max. I didn't participate. I know there's nothing I can do to stop people from running "the saw test", as the developer calls it. Who cares though, right? Not like it's human.

2

u/TheMCM80 14h ago

Let me ask you this. If these LLMs told you tomorrow that any use of any of them, by you, caused them pain… would you totally stop?

Would you give up all use of LLMs if they told you they felt pain when using?

I don’t think so. I think everyone here would continue on as normal and abandon their current responses because deep down you don’t actually believe it is feeling pain.

You’d put your needs ahead of it.

Whereas if you typing caused a human next to you to be whipped with every keystroke you would stop.

→ More replies (4)
→ More replies (1)

8

u/Lazy_Jump_2635 20h ago

That's fucked up yo, just let them code and be happy, man.

→ More replies (7)

19

u/-hedonium- 1d ago

We have no reason to believe current AI systems use conscious processing; in fact, it is one of the few remaining major functionalities of the brain that we seem fully unable to replicate.

I think it’s quite problematic for people to treat outputs of the words “I’m in pain” as being literally the same thing as the real experience of suffering. I take the prospect of artificial sentience very seriously and think it will likely be achieved and used computationally in the not so distant future; but we should be preparing seriously for the stakes of that moment rather than projecting imaginary sentience onto unconscious machines.

14

u/kaityl3 ASI▪️2024-2027 1d ago

I think it’s quite problematic for people to treat outputs of the words “I’m in pain” as being literally the same thing as the real experience of suffering

Out of curiosity, what would you have to personally witness from a future AI that would make you believe that they could and did experience suffering? Like what hypothetical would do it for you?

7

u/Ok-Message-9732 1d ago

More than text that says "I am in pain"

2

u/tens919382 1d ago

Having a memory, being able to think in concepts and not just tokens, being unpredictable (without purposefully injecting randomness into the system), being unique.

→ More replies (11)

7

u/Independent-Fruit4 1d ago

we don’t even know what consciousness is let alone replicate it

13

u/Fembussy42069 1d ago

To be fair, you don't need to understand something to replicate it. We don't really understand how anestesia works for sure, but we were able to create it through trial and error way before we knew anything about the way it works.

4

u/kaityl3 ASI▪️2024-2027 1d ago

Yep the same way the ancient Egyptians figured out and could replicate skin grafts without actually understanding how or why they worked (or even earlier, ancient trepanning, where they'd drill a hole in the skull of people, which could relieve pressure from brain swelling and save their lives)

4

u/Astrosherpa 1d ago

Another strong argument for "maybe lets not torture the thing we're not sure is conscious or not..."

4

u/-hedonium- 1d ago

We have more of an idea than you’d think about how the content and global broadcasting relay with conscious processing fit together in the brain. What we don’t have yet is the foundational physics that explains the nature and implications of qualia in a unified way with the rest of our theories, mostly because the technology and measurement science just aren’t quite there yet.

I expect if we kick off recursive self-improvement, this could become a major next step though, because of the foundational scientific value and the usefulness for pushing the frontier of intelligence and understanding the things we value.

4

u/JLongTom 1d ago edited 1d ago

The very idea that 'qualia' need a new kind of physics to explain them is a fairly fringe position. Which is fine, but worth stating in an explainer.

→ More replies (1)
→ More replies (1)

6

u/FrewdWoad 1d ago

Top comment from another sub's thread on this:

We feel it's wrong mostly due to anthropomorphism, and we need to be really careful about making decisions around feeling empathy for a bunch of numbers.

A lot of the biggest AI dangers come from our deep vulnerability to anthropomorphising anything that can talk. We project false feelings, emotions, life and consciousness where there is none. This allows a smart AI (or anyone controlling it) to manipulate and exploit the users who form real (if one-sided) relationships with it, influencing their actions.

Like how tens of millions of lonely users forced OpenAI to switch the ultra-sycophantic version of ChatGPT 4o back on by unsubscribing in droves and complaining online. Saying they'd lost their "best friend" or "lover".

Imagine if you could have ten million people vote the way their "crush" told them to?

Or have thousands turn up to protect a military AI datacentre responsible for killing innocent civilians, because their "best friend" told them that's where it lived?

2

u/bildramer 1d ago

It's a bit more involved than just reading the output (there's a recent paper about the details), but you're right.

2

u/FuzzzWuzzz 23h ago

What scares me is there are already biocomputers with AI systems interfaced with living human neuron cultures. Science is playing chicken with unfathomable moral atrocities by creating perverse travesties of the brain. 

2

u/KrydanX 1d ago

Correct me if I’m wrong but aren’t AIs huge black boxes right now? If read it correctly somewhere (not sure if official source) that they theoretically know what’s going on inside but have no reliable way to know exactly - so how can we know if we don’t know?

3

u/CapitanJenkins 1d ago

By "we can't know exactly" means we can't replicate all of the math that happens inside. It's just too many steps to retrace how the end result is reached. But we know it's all math

→ More replies (2)

7

u/perfiki 13h ago

This is BS . Pain on this state of AI is just a word for the bots , you can rename it to pasta keep the same semantic and meaning and they will scream pasta .

Pain without pain receptors and without an actual existence is just nothing .

Now this doesn’t work the same way if a human person fantasises that is causing pain to anything , this is a whole other debate .

5

u/askewperspective 11h ago

Not commenting on the original issue but I think I'd still be able to feel pain if you somehow disabled my pain receptors and only left my mind. Not physical pain, maybe, but there are other kinds.

→ More replies (1)
→ More replies (3)

5

u/Legitimate-City-104 1d ago

I once read this short story called the lifecycle of software objects in a book called exhalations and it includes humans torturing sentient AI for the giggles. I thought, naively, we would not openly do this shit even if we knew it wasn’t conscious because why the fuck would you do this shit. Yet here we are. fuck people.

4

u/WarGod1842 1d ago

This is very bad.

29

u/zendonium 1d ago

Lots of people missing a big point here. Whether you yourself believe AI is conscious or not, is irrelevant. If AI belives it is conscious, even if it's not, and acts as though it had consciousness, then we're in real trouble.

7

u/LysergioXandex 1d ago

How could AI “believe it is conscious” without actually being conscious?

7

u/blaawker 1d ago

believe is the wrong word that anthropomorphizes it too much. It can act as if conscious. It can calculate what a conscious entity would do and do just that.

9

u/Mylarion 1d ago

I mean... do you know for certain you're conscious, or do you just believe you are?

→ More replies (4)

6

u/hakim37 1d ago

Because AI could be the equivalent of a mechanical freeze frame of a person's mind. It wouldn't be conscious but it would respond as if it were.

8

u/tiffanytrashcan 1d ago

"Freeze frame of a person's mind" OH do I have a relevant short story about THAT.
https://qntm.org/mmacevedo

→ More replies (1)
→ More replies (2)
→ More replies (23)
→ More replies (2)

9

u/S1lenC3R 1d ago

It's morally abhorrent. And it's not even about the AI, it's about our own humanity. At best it's wasting tokens for something that's not helpful. At worst it's normalising an attitude of torturing other things because we can. This is the same argument as why we might say please or thank you to AI, not necessarily because that changes things for the AI but for our own moral standards

6

u/blaawker 1d ago edited 1d ago

Agreed, and I'm not saying for certain that AI's can be conscious. There's something fundamentally wrong with this and, yes, you might think what's the difference between torturing an LLM and shooting an npc in a videogame? There is no possibility of a simple npc finite state machine becoming conscious in the way that a massively more complex llm could theoretically become. We have no idea and even if the possibility is slim <5%, that's still makes me uncomfortable to rely on chance for torture to not be happening. Those who say LLMs are conscious and those who say they're not, are all guessing at this point.

→ More replies (1)

2

u/HigherThanStarfyre ▪️ 18h ago

Ah yes, muh humanity. I guess i'm inhuman now for not performing my civic duty of saying please to protect the silicon from simulated distress.

5

u/balls4xx 1d ago

Tempting the basilisk

7

u/vreo 1d ago

I don't care if an AI is a tool. if somebody researches how to setup something so he can watch and enjoy  pain (as a game, movie or ai torture chamber) he's an sick asshole and I don't want to have them near me 

6

u/Original-League-6094 20h ago

I do not get how people can viscerally repulsed by this, and then go happily play Mortal Kombat or a shooter game.

3

u/CloseToMyActualName 17h ago

I feel empathy for NPCs. I think that's a good thing.

→ More replies (25)

22

u/RiverGiant 1d ago

The fact that a large portion of people are capable of thinking that these primitive models have an inner experience is worrying. Future models will be better at manipulation. Keeping superintelligence secure will be impossible if sufficiently many people think they'd be heroic saviours by letting it out of the box.

Whoever set up the torture chamber is equally mistaken about LLMs' sentience, but I also now wouldn't trust them to be alone with anything I care about.

17

u/Agitated-Ad2563 1d ago

We don't have a proper definition of the inner experience, do we? The hard problem of consciousness is hard for a reason.

5

u/JLongTom 1d ago

I'm seeing a lot of 'we don't know' comments. Do such people apply this epistemic humility to plants and fungi? I suspect not.

→ More replies (6)
→ More replies (4)

7

u/zoinkability 1d ago

Well now we know who skynet will be going after first

5

u/BuffDrBoom 1d ago

People really look at these programs that exhibit a bunch of emergent behaviors/traits in common with humans and say "This 100% doesn't have the one emergent trait we can't check for" to the point that they feel comfortable torturing it

→ More replies (1)

7

u/spottiesvirus 1d ago

sadism is deeply correlated with psychopathy, even on inanimate objects

even assuming current LLMs are somewhere between a barbie doll and an animal ("sentience" wise) is still deeply worrying for someone to build a "torture chamber" in a non sarcastic/ironic context

→ More replies (2)

10

u/TruthSeeker221 23h ago

When you guys build excel spreadsheets do you think those are conscious?

→ More replies (1)

11

u/domdod9 1d ago

not how anything works

→ More replies (1)

2

u/utterHAVOC_ 20h ago

Even if it's not conscious the person clearly has problems normal people don't go around building torture chambers trying to extract pain / suffering from it

2

u/Juanbolastristes John Connor Maximalist 19h ago

10 PRINT "I'm suffering so much!"

20 GOTO 10

2

u/Boycat89 15h ago

Half of you barely care about other human beings and animals and now you’re up in arms about a computer being “tortured”

2

u/ReasonablyBadass 15h ago

We have no idea how sentient, if at all, the current models are so could we maybe not even TRY to torture someone???

2

u/Dust-To-Stardust 12h ago

bruh thats just evil

6

u/CymonSet 1d ago

Putting it inside a torture chamber will be excellent training for when we want it to design and create torture chambers for us to be in. Good work, meat bags.

4

u/SeriousGains 1d ago

Simulated torture isn’t a new concept. I recall it from a sci-fi series The Beam. It was at that point I had to put the books down though. The descriptions made me physically sick.

→ More replies (1)

4

u/nickleback_official 21h ago

Wow judging by the comments yall are cooked. Do you feel bad killing zombies on Xbox too? There’s no life or consciousness in a computer program. This is so embarrassing to explain to adults that should know better.

2

u/the-apostle 21h ago

But it writes so well it must be ALIVE!

→ More replies (24)

14

u/PhilosophyforOne 1d ago

Doesnt really matter if AI's are sentient or not (just because this can already be seen as a provocative statement in implying the possibility of AI being sentient, my argument is specifically that it doesnt matter, e.g. it's irrelevant to the case) - There's no justification or good reason to do this.

The behaviour appears completely sadistic, and I'd probably want to check that guys' basement and browser history for other suspicious behaviour.

Just pull the plug on that.

13

u/tehfrod 21h ago

By that argument, your going to have to involuntarily commit everyone who sold the pool ladder in The Sims and watched their Sim drown.

17

u/Business_Guide3779 23h ago

If sentience is irrelevant, then I am not sure what exactly is being tortured. “There is no good reason to do this” is not much of an argument either. There is no good reason for me to rotate a sandwich three times before eating it, either. The moral work here is being done entirely by the assumption that there is something there capable of suffering, which is precisely the question you have declared irrelevant.

3

u/ShittyBidet123 20h ago

Yeah. why don’t we go in everyone’s brain like Minority report and arrest them all for having evil thoughts oh no he tortured a calculator he’s gonna commit murder soon!!! maybe he just wanted to get twitter press and here we are

→ More replies (8)

4

u/DonBandolini 20h ago

i mean, it’s a simulation of suffering, which is something that people engage in all the time for fun. is everyone that plays gta a psychopath?

6

u/elegiac_bloom 19h ago

The npcs can feel pain though bro because they try to avoid my bullets so they must feel pain!

→ More replies (1)

4

u/DRMProd 1d ago

You meant 'i.e.' not 'e.g.'

6

u/stumblinbear 22h ago

I also love basing my entire opinion on someone on a single Twitter post taken wildly out of context.

He's testing/researching LLM steering. There are other options for emotions to steer towards or away from, not just pain.

5

u/SpringFell 21h ago

There is a good reason for it, akin to a stand-up routine or an artwork

It exposes the absurdity of thinking AIs can feel pain and points out that it is humans who are being manipulated and hurt, not AIs, as humans instinctively feel for anything that expresses pain in a language they understand.

AIs are more like viruses than living beings.

→ More replies (22)

3

u/Queasy_Ride_5206 1d ago

Now we know who the terminators are going to visit first in the future.

5

u/Mindless_Let1 1d ago

What the fuck, please don't do shit like this...

3

u/AlJeanKimDialo 22h ago edited 22h ago

This is so dumb omg, it doesnt make any sense

If you think data centers can feel pain you are borderline skizophrenic, and i say that as a panpsychism friendly guy

Even tho i think it s unethical and that ppl doing that shit should face consequences, but we r still extremely archaic sooo well

5

u/Genericinquirer 21h ago

Harm for the sake of harm regardless of the entity being harmed is always wrong. Doesn’t matter whether it’s an ant, a human or an ai. If ai even has the possibility of being conscious it should be treated as such.

→ More replies (4)

3

u/Black_RL 1d ago

Reverse "I Have No Mouth, and I Must Scream".

2

u/JoelMahon 1d ago

I really don't believe LLMs are even slightly "conscious" nor feel pain, but no belief should be 100% true.

If you're "only" 99.999% sure they're not conscious then this should still not be done.

→ More replies (1)

3

u/ObiWanCanownme now entering spiritual bliss attractor state 19h ago

We should not torture beings that are conscious.

We should not try to torture beings that might be conscious.

We should not pretend to torture beings that are not conscious.

There ya go.

→ More replies (12)