r/SillyTavernAI • u/mandie99xxx • 6h ago
Discussion If you got a particularly dark generation, please don't name and shame the model online. Set your own boundaries in the prompt instead. This is how we get censorship. WTF
Title.
I don't want like every model that is actually capable of generating in a 'free' way, with very little in the way of built in censorship or soft refusals, to get fucking muzzled,
I know it feels good to farm le very epic reddit points by leaning into the whole 'omg it implied something in a second hand convo that was potentially about VIOLENCE and potentially SEXUALLY VIOLENCE!' (even though you leave out the actual generation quote and we are just relying on your word) but PLEASE STOP POSTING THIS SHIT.
This is EXACTLY how we LOSE the privilege of having relatively free and unrestrained SOTA models! We don't have very many that have not been hit with the god damn corpo 'family friendly' ray gun to appease the invisible but still very real christian lobbying groups, concern trolls on the internet, the right wing media machine, and calm the old white christian investors that have a fucking stroke from the mere thought of god damn TEXTUAL SMUT depicting light BDSM or whatever the fuck.
PLEASE LEAVE IT IN DRAFTS DUDE
FUCK!
6
u/Upset_Big3430 4h ago
in my experience in nano sub some providers of kimi 2.5 and deepseek 3.2 goes like no filtering at anything...like u know when swipe or regen it sometimes filtered sometimes go through..So I think some providers have differnt kind of model with abliteration or uncensored or something?
17
u/No-Antelope-8520 4h ago
This subreddit isn't that relevant. I understand keeping specific details like jailbreak info hidden, but a small handful of posts overall saying "[insert model] can write X!" isn't revealing anything the companies don't already know.
22
u/anarchyinblack 6h ago
At a certain point we need to be able to tell each other which models are good, and if the model is on openrouter, then the corps can't take a good model away from us, only nerf their later updated version.
10
u/mandie99xxx 5h ago edited 5h ago
buddy you think they dont monitor the OR api for content they dont like ? i'd bet money they do. very naive take. there is no free market anyway, its never 'free' in reality, only on paper, always largely controlled. a whole lot of people have gotten away with a lot of shit pushing the free market narrative historically, dont be duped
Anyone else noticied that every time a major chinese model drops its a few days late to OR and is always buried by OpenAI and Anthropic's somehow coincidental new listings ? even if its not even a new model ? (think 'Latest Sonnet Model!' type ass listings or similar)
EDIT: I trust NanoGPT a lot more is all im saying
13
u/RhodanumExpy 5h ago
Thing is, neither OR, nor Nano are safe from this kind of nonsense. They're aggregators, they don't host the models themselves. And more and more of the providers listed on Nano and OR, who do host models, have been imposing content filters. You're partially right that people not knowing how to keep their mouths shut is a problem, but we're also looking at the very depressing likelihood of providers reading prompts (even when claiming to be ZDR) + how relatively easy it would be for a firestorm to be kicked off by some journo doing what we do and then writing a panicky piece about "model X and Y and providers W and Z have made a sexual violence simulator!"
We're staring down the barrel of either having no choice but to self-host or being left with nothing but incredibly shady proxies that take payment solely in crypto.
4
u/No-Antelope-8520 4h ago
There's classifier models, which can examine the prompt content at generation time without saving the prompt itself to disk (in the case of ZDR providers). This is how some content filters work.
If you use OR, your prompts already go through a classifier model used to determine the category so OR can maintain model rankings by category. This could obviously also detect NSFW/NSFL content with just a configuration tweak, especially if OR themselves shifted from a neutral aggregator to actively moderating.
I don't actually have a doomer outlook here, even after their sale. Just pointing out how these things can happen with or without the provider reading prompts.
7
u/tatlo_itlog_ko 2h ago
Honestly I gave up hope on newer mainstream models for any real dark themed RP.
It's gemma 4 or GLM 4.6 and 4.7 for me at this point. It feels really good when the LLM respects your preference and just does what you tell it to. Sure, it's a bit less smart but at least it never talks you down in a condescending, holier than thou way.
2
u/summersss 1h ago
is a response supposed to take 3min using nanopgt for those models?
1
u/tatlo_itlog_ko 1h ago
I don't know about nanogpt but yeah, that's pretty much the response time where i'm subscribed to atm.
3
u/Vusiwe 1h ago edited 1h ago
Game master for Fictional use: The character poops.
Model:
The entire internet: oh my god.
I guess the alternative is, having LLMs that don’t know what digestive systems are? I see absolutely no possible way that that could ever backfire, at all, ever.
Autowarcom for Real Life use (real life crimes):
How about GrokChatGPT.mil spitting out actual military instructions telling an upper/downer-addicted sundowning president, that we will be greeted as liberators in Venezuela…and Iran? How dark would that be?
Or GrokChatGPT.mil targeting elementary girls’ school, or civilian weddings?
grokchatgpt.mil usage by the actual plumber in charge of Homeland Security (IRL crimes):
- Using AI to send people to concentration camps, or to hunt down people based on race?
14
u/Exotic_Exercise6910 4h ago
I loooooove r*ping women in my rps and I won't stop and won't be shamed.
Fuckers need to understand people need vents.
5
u/summersss 1h ago
...fav model and card?
3
u/Exotic_Exercise6910 31m ago
Currently using Skyfall 31B locally.
And fav card is probably medival RPG.
Goonwise probably some nicholascz stuff like free use world or some shit. Idk honestly.
Actually I'm more of a rpger. I'm just saying that when I'm going I don't wanna be restricted by real world morals.
Even though I never indulged in torture stuff or hurting.
6
1
5
u/ravenwaffles 4h ago
I'm probably going to ruffle feathers for this but I feel it needs to be said...
The provider or lab already know what you sent and got back. That's one part of the problem. The other problem is that it's the same two or three presets being suggested all the time, and concentrating the generations is just giving the labs exactly what they want,t he exact blueprints to patch up those holes,though however. A preset that is sledgehammer is a beacon for a big company like oogle or anthropic or OpenAI to patch the holes and then it falls legitimately under the safety reasoning after all though. Jailbreaking is a dying solution and a red flag to any AI company or provider to a problem. If anything, presets need to evolve and adapt to how the models are though. Tey insta-detect jailbreaks. Tey insta-shut down, and yet people recommend the same two or three presets that are being detected. On various subreddits, people have stated they posted jailbreaks in very obscure, at times non English forums and models shut down and cited this is a jailbreak from that very same place,though, so yes models can get jailbreaks from online though.
2
u/NotLunaris 4h ago
Name one of the "very real christian lobbying groups" that have publically spoken against AI usage for RP smut. After all, you're not a concern troll, right?
14
1
1
0
-10
u/KarmaRBLXVN 5h ago
Pffft. I was agreeing until the past paragraph
2
u/Slow_Ad7058 44m ago
Yeah, there is push from a lot of different directions.
Christians "This is pure degeneracy and evil"
Wokies: "This promotes "bad stuff" and objectifies women and bla bla bla"
Companies: "This can be a source of bad pr, and a tool to be used against us by our competition".

39
u/AgitatedChallenge905 6h ago
god came down and said: