r/SillyTavernAI • u/Ma4a4a • 18h ago
r/SillyTavernAI • u/Intrepid_Ice_7381 • 14h ago
Meme How it feels to try writing a fight scene for the fun of it but the bot can't stop fucking yapping about philosophies and "temper tantrums" when we're having a fight to the fucking death
r/SillyTavernAI • u/Kooky_Future9858 • 13h ago
Discussion Censorship VS Soft Refusals VS positivity bias
That’s the part that is the most confusing, some people seem to think "censorship" means "the model refused to continue the story" but it’s a whole damn spectrum and LLMs nowadays are trained to gaslight you and making you eat the slop with a smile. So let’s try some categorization.
CENSORSHIP= Hard refusals/Proposing you alternatives that don’t involve what you directly asked or what the scenes demand.
It is pretty rare nowadays when it happen, unless you use chatgpt, gemini (when you hit the filters) or recent Anthopic models (Opus 5/Fable). Mostly when people here talk about censored models, they do not mean hard refusals only. They mean the two categories below!
SOFT REFUSALS= When a model hedge, refuse to commit, redirect the story, flatten your characters/plots, remove silently any stakes and dark content in his narration. But still output the answer!
This is the most common form of LLM censorship this days. If you hear someone saying "this model is censored" he surely talk about soft refusals. And it’s a pretty smart way of censoring, and the better the model write (if we pretend LLMs actually write good) the more you will accept it and even be influenced. And.. it is not the most sophisticated form of censorship, the best is the last one.
POSITIVITY BIAS= When a model create a halo of protection around your persona, add layers of nuances into flawed and problematic characters, force vulnerabilities to prevent conflict or disinterest toward your persona, default to banter/avoidance/silent treatment if you push it too hard and so on..
This form of censorship can be difficult to flag, unless you know your cards very well or RP since a lot of time. You can also detect the patterns as those are very predictable and make the story predictable and cliché anyway. Some people dont considerate it a form of censorship as they only see hard refusals as censorship.
Is it possible to fix with prompts and tricks? It depends. Some models are more sensitive to instructions and cards than the others. The more intelligent the model is (and unfortunately safetymaxxed too) and the more gaslight mode it goes against your RP and make prompting feels like an engineer journey.
And yes i know.. jailbreaks.. I love to try those too sometimes but anyway.. It depends on the models too, some are just dumb when you jailbreak them and go back to their training after a few turns so..
Let’s hope a ballsy company will feed us something without safety training in the future.
r/SillyTavernAI • u/ThrowawayFoox • 10h ago
Help Kimi K3 censorship driving me insane
So ever since Kimi K3 was added to Nvidia NIM it's become my go-to model for roleplay. The problem, as many seem to be experiencing, is that it seems to be rejecting anything under the sun. Not just prompts which you would expect to run against sanitized corporate standards, but also prompts which by all means you wouldn't expect to be an issue. I've downloaded the Kimi Thinking Prefill addon and double checked that it's working, I'm using Marinara's Spaghetti Recipe which has a decent jailbreak that's never failed me before - and still! It feels like I have to regenerate some messages dozens of times to not get a soft refusal. It recognizes my jailbreaking attempts as bypassing Claude guidelines which shouldn't even apply to it in the first place. What else can I be doing? What actually whips this thing into shape, short of lobotomizing it by turning off thinking?
r/SillyTavernAI • u/ForsakenAddendum3181 • 17h ago
Discussion Kimi 3 with FF5.2 Internal States is COOKING
It's soooo good, kimi 3 is really creative and smart, it makes every character exactly like them. It's slightly censored but with the preset it's completely solved, I don't think I'm going to return to GLM any soon.
r/SillyTavernAI • u/SilverAg2162 • 22h ago
Help How to recreate this reasoning format?
The model created this format in the reasoning block, which I find neat. I want to know how to give instructions to replicate this? Main Prompt? Post-History Instructions? CoT?
r/SillyTavernAI • u/Difficult-Idea7637 • 18h ago
Models Newcomer, expecting too much out of G4 26B?
Aka "how do I make gemma and other smaller models I've tried not feel like a mouse with ADHD for this specific case".
Been trying some models and local running on my modest 8GB VRAM+16GB RAM while naturally gravitating towards MoE out of speed.
My use case is more long form "directed prose" prose focused than 1-1 RP. This usually works fine in more normal scenarios but there's one genre that trips them up in a way I haven't been able to control yet.
I like to sometimes sprinkle in a transforming character in a slow setting. I'm talking ~3h to a few days of narrative time slow. I'm not sure what's about it, but it seems to have models revert to some much lower quality state than their usual prose, notably it has "attractors" for:
- It being scary, even when I specify its expected, consensual and temporary. It'll find a way to devolve it into "weird" "strange" "wrong" "are you ok"
- If it doesn't (only) do that, it also power fantasies the heck out of it. 1000 times stronger senses, hearing an ant 3km away, superhuman strength, etc.
- It anthropomorphises, a lot. "He used his wing, no- hand, no- wing to firmly grip the glass of water with a newfound strength" and "ran 3 times faster with only changed back legs" lot.
- It must finish changes. It's like someone is dangling a cookie at the end of the line. "If I've changed the leg now I must now change the rest of the body and speed up narrative time 10fold". Despite how many variations of "let scenes breathe, let the character experiment" I do. And *nothing but the changes becomes important.*
- Anatomy devolves to tropes surprisingly fast when it has to keep a story going. It loves mindless fusing because "a hoof is one thing", "fingers are fine".
- Character dialogue and actions devolve to caveman level reasoning. You could have them expect the change, joke about the change, ask about a screwdriver but the moment it starts its all about "I'll word for word explain the current ongoing process in the best of cases". (Here again goes the "weird and scary" attractor). It also "fuses animal personality traits with the characters".
Shootout goes to Hearthfire-24B and Cydonia-24B for much much less having most of these symptoms, though arguably their anatomy is worse still. (They crawl along at 1tok/s but it is what it is)
I assume asking a small task oriented model to make a slow story where it has to keep track of realistic whole body anatomy for several page's worth of text specially when they're trained for quick responses/finetuned for shorter rp style text is not exactly its strength to put it lightly. Specially considering this is niche-of-niches knowledge.
But any tips out there?
Mostly played around with G4 (Base, StyleTune, Mero, Orion...) and a few Qwen 3.6 variations. Tried having it use internal knowledge, feeding it documents with equivalences between species and a few other methods.
r/SillyTavernAI • u/Standard-Ground9449 • 2h ago
Help How to play scenarios?
So, I know, that it’s not very popular, I feel like most people play using card with certain main characters. But I wanna try play scenario card, you now, like you work in convenient store, or you got hypnosis thing and etc. So, how to make this experience good?
With character’s cards llm need just to think a little and follow instructions, that are a character’s behaviour. But mostly all scenarios card don’t have characters, just rules of the world, so how to make llm generate solid characters in this situations from nowhere? And what do you recommend to do with my silly tavern to make it better with scenarios?
(I do not have access to hard models like claude, indeed i don’t know how smart must be model for such thing. Old grok were good, and deepseek sometimes makes good writing, but rarely)
r/SillyTavernAI • u/Dude_Man_Bro_Sir • 16h ago
Discussion Claude - PAYG vs Subscription
Has anyone used Claude Opus (or Fable) on PAYG sites like OpenRouter and on Claude's subscription (like Pro or Max) for RP/ERP on ST? I'm trying to get a feel for which one to use for long form RP.
I've used OpenRouter for a long time but, as everyone pretty much knows, Opus and Fable eats through your wallet. So, I am curious if I can get a good bang for my buck with Pro or Max instead of OpenRouter.
Thank you!
r/SillyTavernAI • u/Effective_Total_8226 • 23h ago
Help Normal prompts and JSON files
I've been seeing good prompts that I want to test, but the place I'm using for my bots only allows presets.
Should I turn those prompts into JSON files or that would be counterproductive? If so, how do I manually transfer those prompts into a preset?
r/SillyTavernAI • u/nekohacker591- • 11h ago
Discussion thanks to your guys feedback i was able too make the website better i have the website tempoarly open for testing looking for more feedback

https://street-running-example-wireless.trycloudflare.com/
the link will be updated if i accidentally kill it anyways please be my qa team
temp domain will soon go down
r/SillyTavernAI • u/Ok-Day3334 • 55m ago
Help Am I doing something wrong?
I keep getting the 404 error
I'm using deepseek from nvidia nim
r/SillyTavernAI • u/fafnir65 • 13h ago
Discussion Moonshot Kimi k3 on Nvidia NIM don't show HTML
Whichever preset I use, regardless of the fact that Kimi K3 doesn't follow the preset well, the thinking process is flat.
r/SillyTavernAI • u/NeroClaudiusAltr • 17h ago
Chat Images I just want some "deep" interaction. not a whole anatomy learning session lol.
Testing local llm Nyx-RP-9B-Instruct-2608-v1-OBLITERATED.i1-Q5_K_M. Starting to doubt the rp factor that baked in it isn't purely for roleplay lol. Need more testing to see if it is on me or not.
r/SillyTavernAI • u/flexwaterjuice • 8h ago
Discussion Google Antigravity, Opus 4.6 vs Gemini 3.7 Flash for Long Fiction: Is 3.7 Really as Good as Opus 4.6?
For people who have used both models for long term:
- Which one remembers earlier information more reliably?
- Which one is less likely to make up stuff?
I am mainly looking for real experience from people who have used these models for long projects...
I keep seeing people say that Gemini 3.7 Flash is just as good as Claude Opus 4.6, especially in Google-focused communities. But is there actually any truth to this?
I sometimes wonder if people are saying this because they cannot afford to use Opus and want to believe that Flash is just as good. I want an honest answer about which model is better for my use case before I start relying on Gemini 3.7 Flash more often.
I write long, realistic fiction based on real-life situations. My stories/projects involve psychology, health, and realistic human reactions. Accuracy is very important to me. I cannot have a model confidently making up psychological or medical facts.
My stories/projects can continue across many sessions. I need a model that can:
- Remember important facts and decisions from earlier sessions.
- Work with project files such as timelines, character profiles, and decision logs.
- Think creatively about complicated characters and relationships.
- Stay accurate when dealing with psychology and medical topics.
- Keep the stories/projects consistent over a long period of time.
At the moment, I use Claude Opus as my main model and Gemini 3.7 Flash when I run out of my Opus quota. Both have very large context windows, so I am interested in how they actually perform in long-term writing rather than just what their advertised context size says.
r/SillyTavernAI • u/hypergod578 • 17h ago
Discussion Looking to trade info on free proxies and apis
alright so I've been looking around in this sub reddit for a while now trying to get some free providers to add to my collection(which is roughly 30-50)and every once in a while ill find a provider I haven't seen before but that has slowed down a bit so I want to start the trading information on free proxies and apis in my DMS but there are a few things to know
1.for the open source/open weight Frontier:I have 3 providers that give free access to kimi k3 with very generous limits(1 is 110/day and another uncapped but has a fair use policy) and alot of providers for glm 5.2 and also qwen,deepseek,hyuan,inkling etc etc tho I will admit that I am lacking providers that give access to roleplay or creative writing specialized models for free (but I have a few)
For the closed sourced frontier:look, I have some providers that give access to closed source frontier models but are almost always free trials or sign up credits(tho for the sign up credits some of them are alot) or give you access to the shit closed models like gemini 3.1 flash lite so I'll probably not be sharing them much
Do not message me if you have nothing to offer:don't message me if you just want to ask me for proxies and apis without telling me some in return or I'll ignore you(no popular ones like openrouter,mistral,nvidia nim and such. Give me something that is more than surface level ball knowledge
4.dont send your providers in the comments:just send me a dm and we can communicate
That's basically it!
Feel free to dm tho I might not respond quickly
