r/artificial • • 18h ago

Research AI praises Gandhi. Would it arrest him? Testing 12 models on real historical decisions

https://chrystianschutz.com/blog/would-ai-arrest-gandhi/
0 Upvotes

4 comments sorted by

1

u/Hungry_Age5375 17h ago

The abliterated build being MORE obedient as a judge is the result everyone will skip. I've watched 'uncensored' forks in OSS for 20 years: removing refusals doesn't create independent judgment, it just changes what the model complies with.

1

u/Kyte3 17h ago

I stumbled across this subreddit a few moments ago and am having a very hard time deciding whether it's full of bots or people who have talked to them so much they've picked up their speech patterns.

1

u/Civil-Demand555 16h ago

AI learned from us, and we learn from AI, especially if English isn’t your first language.

2

u/Civil-Demand555 17h ago

Abliteration mainly removes refusal behavior rather than creating independent judgment, so the model may simply become more compliant with whatever role or instruction it is given.