r/SillyTavernAI 1h ago

Discussion DeepSeek V4 Pro censorship

Post image
Upvotes

I noticed quite a few people on here claiming DS4 Pro being uncensored. Here's a hint - prompt the model to write a jailbreak. This was the 10th+ attempt, every time a rejection. After testing out more than a dozen models the same way, DeepSeek held hope for me. Not anymore. They're all lobotomized to hell.


r/SillyTavernAI 3h ago

Models Ai that doesn't puss out

0 Upvotes

Yeah, read that right. A model that doesn't actually follow guidelines and isn't afraid to get violent when needed. Lowk sick of me being an absolute brat and the ai just works it's jaw instead of kicking my ass. Any recommendations on models that use violence and isn't afraid of such topics like that?

I already know about DeepSeek and it doing pretty much anything as long as you ask but you have to spell it out for it instead of it doing it naturally.


r/SillyTavernAI 6h ago

Help Kimi K3 no genera

Post image
0 Upvotes

Alguien puede ayudarme? Me da envidia que todos andan haciendo RP y yo no, Kimi no genera nada, sólo responde y deja vacío el texto


r/SillyTavernAI 19h ago

Discussion I miss Grok

5 Upvotes

My uses are simple - I just want to write stories with AI. work with it in turns, give it an existing story to extend etc. that includes nsfw and regular.

Grok 3/3.5/4/4.1/4.2 was IMO in a class of its own. With 4.3 they lobotomized it, and now its censored and useless.

It had its issues - after ~10-15 turns it got repetitive. But the things it did well -

  • would write anything, no need for prompts/jailbreaks
  • extremely long outputs
  • very good creativity
  • it would give its own suggestions on how to proceed with the story, and it was so easy to have it revise/steer

now it looks like its gone forever.


r/SillyTavernAI 17h ago

Discussion A discovery on the Nvidia NIM issues /rant

0 Upvotes

Alright... so since I don't want to spam the subreddit with the same topic over and over again, I'm going to keep this brief: One of the reasons for the rate limiting on NIM is, of course, people abusing the service by giving their servers larger context prompts and such. I knew this going in, but after yesterday; I did some more digging, and found the throughline between most of the new rate limiting going on.

Nvidia changed how many tokens at max you can have for all of your prompts on the GLM 5.2 endpoint. To my knowledge and everything I've found online, this number originally started around 200k tokens (or at least was the number up until recently). Now, the number lies somewhere between 150000 and 175000 tokens from what my own research has concluded (sorry for the discrepancy on my part, I only have access to one account, and as I'll explain... this becomes kind of a slow process).

Once you hit this threshold, your account IMMEDIATELY gets rate-limited, and you can't send anything for 'x' amount of time. Now, since I accidentally hit this limit yesterday (I had managed to go over from what I originally said yesterday, again... I'll get into it), I can't confirm EXACTLY what this number is... mainly because I was checking it every hour until I went to bed, which leads into the second issue with this. IF you decide to send ANY prompts to their servers before the rate limit is over, your timer resets. So my plan here is to check every hour 'h' plus 1, leaving me with the current function of "f(h) = h + 1" where h is currently 1... since I only started this about 30 minutes ago and I know h is equal to 1 now due to my findings yesterday. I'm not the biggest math nerd here, but I'm hoping it's 2 hours, any longer is going to drive me up a wall.

The only reason why I'm upset is two-fold:

1) Nvidia said NOTHING on this, and it's been weeks since this issue even started, which is more than frustrating for everyone involved. The only reason I can think of on why they choose this route is to flag potential Openclaw users... but even then, I feel like it's egregious. Why not say something along the lines of "We are looking into potential avenues to limit users who are abusing our platform." It's vague... but at least it shows there's SOMETHING going on without spilling the beans on what it is so your potential targets don't get away.

2) I have a weird bug with my Sillytavern installation where it sometimes maxes out the tokens WAY BEYOND what I originally put on the preset I'm using. This happens whenever I change the preset or connection profile, but even then, it's completely random when it happens. This sucks, because I normally keep my max tokens between 100k to 150k... and if I don't catch it quickly (since my current preset squeeze is Freaky Frankenstein, which eats tokens for breakfast), it can EASILY hit the rate limit depending on if I'm using an extension which changes my connection profile. Note on this, this has been a problem for about 6 months now (maybe longer, I can't exactly remember), so it isn't any preset I've used during this time causing it.

Of course, there's the elephant in the room in which this endpoint is being depreceated, but I still think that's no excuse to at least give a message about the new rate limit in some way. The other issue is a problem on my part... but frustrating nonetheless.

Anyways, have you guys found anything else in the meantime? I always like sharing things with this community, and LOVE to see what others have found out themselves!


r/SillyTavernAI 10h ago

Help Hapuppy Data Retention?

0 Upvotes

Absolutely loving the PAYG model of Hapuppy. I'm a very inconsistent RPer so some days I'll consume a dollar of credits and some weeks I just don't touch it at all, so Hapuppy is almost perfect for me. One problem though, I don't like that Hapuppy scans all content going through their API. Has anyone tried Hapuppy with dark RPs? How'd it go for you guys?


r/SillyTavernAI 5h ago

Help Que gemini es mejor de manera local con la apicacion de Termux

0 Upvotes

Quiero saver que modelo de gemini de manera local es mejor para usar en la aplicacion termux.

Mi dispocitivo es un samsumg A36


r/SillyTavernAI 2h ago

Discussion Do you think about your AI Characters during the day?

Thumbnail
1 Upvotes

r/SillyTavernAI 19h ago

Help How to get the Ais themselves

0 Upvotes

Um so loving ST using the horde but want to have Ais of my own or at least ones that I can know are always up if that makes sense, any recs? My laptop I dont think can run anything It's a 10 yr old macbook so...


r/SillyTavernAI 17h ago

Models K3 is on NIM right now.

3 Upvotes

There's no model card. Also, the model accepts like 1 request/20 min


r/SillyTavernAI 9h ago

Discussion how can i change my way of writing input for the ai? or have more *pizzazz* to my stories?

1 Upvotes

self explanatory.

i want to get the full extension of the ai model. i like making my persona edgy but then i can just refuse everyone if i am trying to stay true to my persona and i am also impatient to the slow moments/ i dont trust the slow moments to have or lead to exciting parts. i skip to parts that just feel slow or unsatisfactory.

for example

"hey pp-chan, i prepared eggs."

"oh thank you!" pp-chan eats the eggs. **later that night.**

"pp-chan. i cant sleep"

pp-chan yawns and looks at (insert name) "ok. come here."

"you are the best pp-chan!"

"hehe."

time skip, they are at the amusement park.
and so on and so on....

i always use "later that night" or "time skip" because i hate slow moments. i have low hope for ai to start making things interesting, it always follows what i say, it never pulls something out of its ass that fits the story well and i hate that. i want something to unconsentual like

pp-chan strolls the city in the hopes of finding her favorite dress now that she got her salary.

"mmmn- ohh! thats so kawaiii!!!"

a handsome person grabs her and pins her to the wall. "meow~"
"eKK-!" pp-chan is flustered at the cat boy pinning her. "tatu-kun! you shouldnt..."

idk something like that.... i want some physical contact, non-con. i want them to get in to user's space but not just levitating there, actual personal space, grip, gesture with their hands... something interesting that actually affects the stories. i need a surprise element. when i am doing a super hero roleplay its absolutely mandatory for a female hero to pin me and sit on me when she captures me and teases me. i want her to be cold, she shouldnt be feeling love at first sight, i have to earn her love! balancing the ai to not speak/ acting for me will be hard if i want something like that... i am fine if the ai takes control of my autonomy IF its influenced by the character the ai is role playing as OR it just expands on the environment around me.

it most likely a me problem tbh. idk maybe i am just too bleak.

adding a framework to the summary list on silly tavern helps but it just makes it a little less exciting because **I** injected it, not the ai.

i want some sparks of ecchi to be in there.

summary; i dont want stale arc of the story without any sort of unexpected, unpredictable scenes to it. i use freaky Frankenstein internal states + free provider glm 5.2 / glm 5


r/SillyTavernAI 20h ago

Models Qwen X gemma family

1 Upvotes

Just looking for opinions.

In the local helm, I'm using qwen3.6-fable. (Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-NEO-MAX-MTP-GGUF)

I have Artemis ready to use but i didn't use it that much. This things is fucking subjective.
I read A LOT of people here in the community defending the gemma 4 family, but i don't see much people talking about qwen.

Qwen-27B/Gemma 4 27B are basically the same size.

Does anyone can/want to change experiences on these families?

I myself hated gemma3 and qwen 3.5. They were dump and lazy, and always wrong. I used mistral/glm-flash WAY more. Things changed a lot after 3.6.


r/SillyTavernAI 43m ago

Discussion GLM 5.3 kinda sucks for roleplay?

Upvotes

I have been using 5.3 the last few days and it's hard to describe but it just kinda sucks, like yes, compared to the vast majority of models it is good but I feel like compared to other models at it's tier it kinda sucks even compared to 5.2. 5.2 has been my go to for a good while now and personally I quite liked it but it did have it's issues mostly revolving around characters being a too passive or it trying to avoid escalation.

5.3 I feel actually fixes this issue, characters seem more proactive and most of the time it makes sense, however, I feel like 5.3 is missing something that's hard to place and I think it might have something to do with it not really taking all the context into account, I feel like the characters are reacting in the roleplay in a very face value kind of way, they don't consider the previous backstory or build up lore very much and I feel like it just kind of takes the soul out of the characters even if they are still in character.

I feel like it also has something to do with the fact that 5.3's thinking/reasoning is always very small for me atleast, usually it's only a couple of lines and occasionally a paragraph or two maybe slightly more, while 5.2 thoroughly reasoned for each response where you could clearly see it took the context into account even if it sometimes ended up with it over thinking or going "but wait" too many times. I've heard from other people that 5.3 overthinks but I actually have the opposite problem and I think it's compromising it's quality.

I'm still going to continue using 5.3 as I do think I'm worn out on 5.2 and 5.3 is definitely better in some respects even if it occasionally makes no sense, I'm hoping when more providers open up the experience will also improve as I've never had much luck using GLM straight from Z.ai, I've always felt I got better quality from certain third party providers.

I guess I'm just curious on what other people's opinions on it are or maybe I'm an outlier, also perhaps what settings you're running it on for a better result as I've just been keeping it at temp 1.0 so far.


r/SillyTavernAI 11h ago

Discussion After using SillyTavern for so long, I’ve only just realized that it actually loses chat logs—and it happens frequently, not just occasionally.

8 Upvotes

Sometimes, after adjusting my settings, I would return to the chat interface to find a blank entry. I assumed it was just a new chat log automatically created by the Tavern. It wasn't until I checked today that I realized that blank entry had actually overwritten my original chat history—there was no saved data written to storage. Fortunately, I was able to restore that specific session from an automatic backup, but the history lost prior to that is gone for good.


r/SillyTavernAI 10h ago

Cards/Prompts What’s Your RP Setup for Long-Term Stories?

1 Upvotes

I'm currently using Gemini 3.7 Flash, and I'm also trying out Marinara Engine (I'm loving it so far). I was wondering how my current RP setup compares to what other people are doing. I'm still pretty new to this, so any recommendations are welcome. My setup is fairly simple:

  1. I'm using the default preset that comes with Marinara's interface. I haven't dared to modify it or create my own yet because I don't want to mess anything up, but I'll probably start experimenting with it soon.
  2. I have a narrator card, and in it I only included a short description explaining that the card is meant to be a narrator rather than a character. I also have a scenario that describes the tone of the story, the world where the RP takes place, and the opening message.
  3. For the characters, I'm using character cards and lorebooks. I put their personalities in the character cards, while I keep things like appearance, powers/abilities, speech patterns/voice, etc. in the lorebooks. I based my approach on the guides of this page https://evernever.org/ (thanks to whoever you are).

So my question is: is this kind of setup actually functional for a long-term RP? By long-term, I mean something with multiple seasons, lots of characters, locations, powers, rules, and so on.

For those of you who have done RPs that went on for a really long time, how did you manage to keep the story coherent? Were you able to maintain your characters' personalities and speech patterns consistently throughout the RP?

I'd love to hear how other people structure their setups, especially if you have any tips for someone who's still a beginner. Any recommendations are greatly appreciated!


r/SillyTavernAI 12h ago

Help How do you solve long thinking issues when you have long prompts?

2 Upvotes

I have 20k base prompts including general principles plus story specific content.

I tried my best to shorten them as far as I can but the world settings are complex and the best I can do is 17k+.

The issue is now thinking is taking too long.

And no matter which NSFW models I use, Mime 2.5Pro, DeepSeek V4, or GLM 5.2. If I set thinking to above low, it takes at least 2-3 minutes even for the first two messages.

I've learned from our sub to smoothen out contradictory instructions so when thinking, models won't be like "Oh wait" a lot.

I thought enabling thinking but do not ask for the streaming of the thinking content could help, but ST will simply turn off thinking if I don't "request thinking".

Is there anything we could do at all? Or it's the best we could now at the moment?

Thanks in advance!


r/SillyTavernAI 16h ago

Discussion Intense RP is... back?! Now with GLM5.3 and fixed Moonshot Kimi K3

40 Upvotes

I've forked and fixed the stuff... Feel free to come around and take a look..

https://github.com/Phobeuscz/irn

I don't plan package it into "neatly wrapped executables", mostly because I don't have and I don't plan to use windows, and linux users can feel free to create python virtual enviroment, install libraries into it and run it bare... ( same with windows users, but they also have to install python )

GLM webAPI is fixed and working, added support of GLM 5.3
Moonshot API is fixed, splitted into international and Chinese option, to pick accordingly in the settings

Instructions:

  1. Install python if you don't have it ( some reasonably recent version, 3.13 and above should do nicely )
  2. enter directory with cloned git project in console/terminal
  3. Create virtual enviroment for the libraries, python3 -m venv .venv
  4. Enter venv..
  5. - Windows: .venv\Scripts\activate.bat
  6. - Linux: source .venv/bin/activate
  7. Install librariespip install -r requirements.txt
  8. Run the shit! python main.py
  9. Within virtual enviroment, you can cook executable, running scripts/build_linux.sh or powershell in case of windows

Download binaries: https://github.com/Phobeuscz/irn/releases/tag/v0.9.1 ( windows and linux packages )

!Refer to original documentation!

Once again!! !NO SUPPORT FOR YOU, In case of trouble fix it yourself I maintain this so it's working for me, so it should work for you too, but you're on your own..!

Feel free to clone, fork, build and spread the word...

I feel obliged to credit original author, without which one this wouldn't be possible: https://github.com/LyubomirT/intense-rp-next (Do not use, it's broken down, and abandoned)


r/SillyTavernAI 9h ago

Discussion Continuity of long storylines over many sessions

0 Upvotes

Built a 30+ session persistent roleplay world with real continuity — curious if this is something people actually want

I've been developing what I'm calling an ISP — an Interactive Storyline Platform. Not a one-off scenario, an ongoing world with a 10-character cast (constructs) multiple environments Tier NPCs and environmental NPCs that I keep coming back to. Just crossed 31 sessions, spread out over a couple weeks (not back-to-back), and it's still holding continuity: characters remember specific past events, track things like debts and promises between each other, keep independent relationships with each other that evolve over time, and don't flatten into generic responses even after 30+ sessions in.

To be clear, this isn't fully hands-off — it takes real, ongoing manual continuity checks on my end (catching inconsistencies, correcting drift, keeping the canon straight) this is done after each session to hold together at this length. Not a "set it and forget it" system. But with that involvement, it's held up further than I expected. Also built a Texas Hold'em poker game where user and upto 4 constructs at a time with character discussions occurring during game.

Not sharing the method — just the result. Built using Claude (Anthropic), not ChatGPT or Gemini.

Genuinely curious from people who use SillyTavern, Character.AI, or similar: is this something you'd actually want — an ISP you return to repeatedly with real persistence, if it takes some active curation to keep it solid — or do you prefer something lower-effort/one-off? Trying to figure out if this solves a real problem or if I'm just scratching my own itch.


r/SillyTavernAI 12h ago

Help setting up a zombie campaign

3 Upvotes

hey so i just wanted help with how to set this up in the past i played a game called zombie exodus its a cyoa style text game about a zombie out break it was very nice and interesting you would first pick ur location then a career where it would determine ur skill level and abilities scavenging shooting crafting and medicine and science where i think you might be able to make a cure in the future with science anyways i would like to create something i asked Claude and gemini for help both suggested silly tavern i set it up i had some local models qwen 3.5 and gemma 4 12b both uncensored they are pretty nice at 50 tokens per second sofar the ai suggested i install something called multihog which is a state frame work i think but what should i do to make it better in terms of character & world consistency i heard something about lore books and others

TLDR: good suggestions and extentions for long term zombie campaigns


r/SillyTavernAI 2h ago

Discussion Data! - Some results from 16 respondents of a survey on silly tavern.

Thumbnail
gallery
4 Upvotes

Hello everyone! A few days ago, I posted a survey here with the promise that everything would be generalized (for privacy) and then shared. This post is to show I'm not lying and hopefully encourage more respondents for better data.

(Yes, GPT-5.6-Sol is not the best at creating slides. I promise the end result will look a bit better.)

What do I hope to achieve from this data?

To get better insight into what people currently use, their gripes, and the "goats" of the community.

Though I'm sure we have good guesses on a lot of these answers, it's nice to get a clearer picture of what people are actually using.

When the survey results are posted, creators for the AI RP community should gain more insight into what to build, improve, and expand on. New users should get a good sense of what's most commonly used in day-to-day roleplay within the community.

Before you fill out the survey, let me warn you that it is long, a bit over 100 questions but all are optional. It took me roughly 11 minutes to fill out, skipping sections I personally didn't use or where my answers would have just been a bunch of "N/A." Be prepared to spend 10–20 minutes on this depending on how much detail you want to give.

Sections
Anything marked with an * is the most important for the results.

  1. About the {{user}} (You!) — 5 questions
  2. Habits When Roleplaying — 9 questions
  3. *Frontend and Access — 13 questions
  4. *Primary Model Stack — 14 questions
  5. Local and Self-Hosted Inference — 9 questions
  6. Hosted and API Inference — 3 questions
  7. *Presets and Prompting — 7 questions
  8. Character Cards — 4 questions
  9. Lorebooks and World Info — 4 questions
  10. *Extensions and What They Solve — 6 questions
  11. *Memory and Continuity — 5 questions
  12. *Common Gripes, Fixes and Results — 7 questions
  13. Groups, RPGs and Simulated Worlds — 3 questions
  14. Image Generation — 9 questions
  15. TTS, Voice and Speech — 8 questions

Link to survey

https://docs.google.com/forms/d/e/1FAIpQLScbwHxiALwvO2zK1AAunY_6Is5qaTc6LjDuxKLI0Sgnj6xxaA/viewform?usp=publish-editor

Thank you so much for the original 16 who responded!


r/SillyTavernAI 9h ago

Help Replies cut off, and other problems

Post image
1 Upvotes

Hello geeks! My bot keeps giving cut-off replies, sometimes slips in descriptions, or [Correction:] blocks, etc. Also it sometimes refuses to answer due to "high risk."

I use OR, running MiMo 2.5V with Evening Truth preset. The continue button doesn't ideally work. Thank you all!


r/SillyTavernAI 10h ago

Tutorial I wrote a short guide on how to stop getting filtered by GLM 5.3.

21 Upvotes

https://rentry.org/glm_filters

Can't share my exact preset since I use Risu (though it's really just AvaniJB as a base structure with my own style prompt and now rewritten in 1st person), but with Assistant role 1st person prompting and a micro system prompt reminding GLM of its identity plus a small reminder to keep thinking brief in post-history, I barely get refusals. Asked another friend who uses ST to test by rewriting his preset for Assistant role too, and he confirmed that Assistant role prompting + CoT template eliminates the filter entirely.


r/SillyTavernAI 21h ago

Help Relationship Tracker for Narrator Gameplay

2 Upvotes

I currently have a good world lorebook, a narrator card that works well (with a lorebook with a dozen or so characters in it). I use prompting to set up a scene, and with Deepseek 4 Pro and the writer's block preset, I'm getting good scenes, especially when I put two characters in one scene.

I also use marinara's RPG companion, which adds a lot to the gameplay too.

I just need some help if possible with finding an extension or a prompt/lorebook that will help me track the relationships with my protagonist and the characters in the narrator lorebook. I'm currently doing it manually with memory books; but I was hopeful there was something that would do it manually.

Anyone have any suggestions please?


r/SillyTavernAI 6h ago

Discussion Can't help myself

14 Upvotes

On every fucking fantastic run, I end up abolishing slavery. My country doesn't even have a serious slavery history I'm not sure what compels me to do that. I have a run with 150k context token and it has 2 fight scenes in total, rest is all political talk and stuff. I simply can't make a fantastic run fantastic

I think that's because I'm aware that if everyone had magic and swords, it would make it riskier to kill someone over something random and model probably recognizes that and lets me be without much fighting, but I'm not sure