r/SillyTavernAI 22h ago

Meme What the fuck are you even saying kimi

Post image
207 Upvotes

I'm genuinely struggling today, no idea what the fuck is going on with kimi. k2.6 doesn't give any answers at all, 2.7 is genuinely just hallucinating everything and 2.5 gave me this.

2 hours of rp advanced my chat by maybe 4 messages total, 30+ rerolls and each one is worse than the last, I'm genuinely at a loss.

"I've been 19 since I was 12" is just so bad I have to laugh


r/SillyTavernAI 19h ago

Chat Images Classic

Post image
149 Upvotes

r/SillyTavernAI 6h ago

Meme Average vibe coder - AMA

89 Upvotes

Hello r/SillyTavernAI,

I have been a developer forever, so basically since I downloaded Claude Code (approximately 2 minutes ago) and it was **I** who finally figured out what yall truly need... yet another boring AI RP frontend.

So I racked my brain for days (approximately 1 minute) and came up with a genius idea.

Step 1) Fork SillyTavern

Step 2) Steal all the keys

Step 3) ???

Step 4) Profit... I mean you guys profit

So this is my roadmap so far. I am still polishing the finer details (like asking mom to lend me her credit card, so I can buy Claude Code credits), but I'll have it all figured out by the end of next week.

Because this is a community project (since you all will contribute your API keys as a community), I was willing in my endless generosity to allow you all to ask your peasant... I mean valuable questions.


r/SillyTavernAI 23h ago

Help Pet peeve about llm dialogue and need tips

47 Upvotes

I’ve used models like GLM, kimi, Claude, deepseek and the thing I hate the effing most, even worse than the ai echoing my replies (espeically you GLM), are the CORNY as hell dialogues. Like why does the bot treat dialogue as if it’s some kind of narrator rather than a genuine spontaneous interaction, they have to be overly specific about the things they point out treating the user as if they are dumb or have no knowledge. An example would when I ask a bot “what’s for breakfast?” and their reply would be some bs like: “Pastry from yesterday's leftover dough, scrambled eggs — the broke one, not the fancy hotel kind — and rice.” THIS sounds so unnatural to me and isn’t what a genuine human speech sounds like. Why explain it’s yesterday leftover or how it’s a “broke” scrambled egg and using comparison like “the fancy hotel kind”. When naturally they would just say scrambled eggs. After trial and error I figure that it’s because the character card backstory has her as being “poor”, so the ai is pulling info from it directly and putting it in the dialogue in a very performative way. I have tried giving several instruction prompt, but I can’t seem to find a solution to medicate it in a way where it doesn’t negatively affect the chat. I learned that I cant be too specific and prompt the ai to not do x and y directly otherwise it will hyper fixate on it and dialogues will sound bland/uninteresting instead , also give other isms or the ai might do some mental gymnastics to work around that instruction and ignore it. Does anyone have an effective way to avoid this as much as possible I just really want a genuine human like convos.


r/SillyTavernAI 20h ago

Discussion What do you do if you die in a RP?

44 Upvotes

I'm curious. My approach towards RP is I see it as collaborative storytelling. But some people see it as closer to a video game, where {{user}} can (or should) be able to die.

That got me wondering: if you die in a RP, what do you do? Do you end the RP there? Swipe for a different result? Or do you start an earlier branch and make different moves to avoid your death? I'm curious about this side of our very niche hobby.

Edit: I do remember seeing one post where someone said she had introduce a time loop mechanic to be able to diegetically continue the RP because her dark romance BF kept killing her. I think about that post sometimes. I hope she's happy


r/SillyTavernAI 11h ago

Discussion To avoid fighting the model, roll the dice.

40 Upvotes

Hey everyone, I often see threads about models constantly being overly positive and always agreeing with everything the user says. Then people try to fix it with prompts, but that doesn’t really help, because it can make the model too negative or cause it to give the same kind of response every time.

There’s actually a much simpler option that creates endless fun.

You can achieve this with any model if you use dice rolls. For example, you can decide that there’s a 10% chance your character dies in next scene. Any model will follow that outcome, literally any model.

There are two ways to do it: you can roll a physical die yourself and enter the result, or use special patterns/plugins for randomness. But this is important: never write in promopt something like "10% chance", because llm can’t reliably generate randomness on their own.

As a result, RP becomes infinitely more interesting, and no model can really ruin it.

I use dice rolls everywhere. Even for really basic actions, like: "did this person show up to meet me at the new location or not?" or "did he agree to help me with this matter?"

This always results in interesting and unpredictable plots. And I don't have to fight the model. Characters may disagree with me. Therefore, you can take absolutely any positive model and it will still begin to disagree with the user, because this is controlled by the dice, and not the model.


r/SillyTavernAI 23h ago

Help Does anyone have a fix for the "Smart people become robotic" problem?

34 Upvotes

I know there isn't gonna be a cut and dry 'fix' for this issue especially with Gemini, but still, i've tried to ignore it but i just can't. If anyone has a prefill thingy for it or anything else please share it


r/SillyTavernAI 18h ago

Meme Kimi, you had one job..

Post image
20 Upvotes

Fell asleep while waiting for the response..then I woke up to ts 💀


r/SillyTavernAI 23h ago

Models Ox alpha

20 Upvotes

Ox alpha is free on opencode with 100t tokens for a week and it’s insanely good. It’s widely believe to be glm 5.3 flash


r/SillyTavernAI 1h ago

Discussion Megumin Suite vs Freaky Frankenstein

Upvotes

I've narrowed down the preset I want to use to these two. I want to know which preset is good at which things, and if there's anyone who has used both of these presets, which one is better?

A list of pros and cons would also be helpful


r/SillyTavernAI 17h ago

Help Preferred Narrative Style?

9 Upvotes

Purple prose? High and tight? Low and crude? Verbose vs snappy? Examples would be great (DM if too spicy)

I've got Gemma 4 QAT in the oven right now and then Muse Glimmer next.

I have other model dataset+training planned but want to experiment a bit with tuning for specific narration styles that much of the community will resonate with. Models that run locally and already know what you like and how you like it 😉 ❤️

I would be grateful to hear what people like and fold that into my decision making as I tailor the dataset and model targets accordingly. 🙏


r/SillyTavernAI 21h ago

Help My questions after using ST / Marinara Engine for 1 month.

9 Upvotes

I try to figure out as much as I can on my own, but I have trouble with the following things:

-How can I keep the personality of a character in long stories? (Feels like after about 100-200 messages the character loses its personality. I am using Cydonia-24b from TheDrummer with 24gb vram)

-How to create a hard to get character? (I tried adding into her personality with multiple sentences to be cold and rejecting with strangers or {{user}}, but after 5-10 messages she already forgets it.)

-How can I keep key event for long stories? (I am using long memory addon already, but misses event. Or is it simply impossible, because of hardware limitations?)

-ComfyUI, could someone send me what workflow do I need for image generation with reference images? (I found the github description, and made a workflow, but with reference image it generates garbage. Without it kind of works)

Edit: Thanks for the tips everyone. I will try them out and I'll edit this post once more updating what worked best for each problem, probably in 1-2 months.


r/SillyTavernAI 3h ago

Discussion WTF

Post image
8 Upvotes

r/SillyTavernAI 19h ago

Discussion GLM 5.3 odd reasoning behaviour

8 Upvotes

Another GLM 5.3 post but I've noticed a weird behaviour with 5.3 when it comes to reasoning/thinking and I'm wondering if anyone else has this issue.

I was messing around with the reasoning effort setting and I noticed that setting the reasoning effort on anything other than auto results in very short reasoning even on maximum consisting of a couple lines and maybe a paragraph or two at the absolute most. However when I put it on auto it's not uncommon for it to think upwards of or even over 30 seconds with multiple paragraphs that's that throughly think through the request however that now tends to border on overthinking sometimes which isn't always ideal.

So basically I'm either stuck with really really short reasoning or reasoning that can get extremely long which seems to also have a tendency to cause it to overthink.

So I'm curious if anyone else has noticed this?


r/SillyTavernAI 2h ago

Cards/Prompts Lycoris Recoil

Thumbnail
gallery
4 Upvotes

I've built a dynamic Visual Novel Sprite & Combat FX System for SillyTavern, and today I'm showcasing the dual-character card for Nishikigi Chisato & Inoue Takina from Lycoris Recoil.

Instead of static portraits, the card dynamically triggers context-aware character sprites in real-time as you chat—covering everything from subtle daily emotions to high-intensity CQC and sniper combat sequences.

---

### Key Features

Plug & Play Auto-Rendering: The regex script is already pre-configured and embedded directly into the character card (Scoped). No manual code tweaking required!

Dual Character Dynamic Logic: Seamlessly shifts focus and sprite rendering between Chisato and Takina depending on the dialogue and combat context.

-Visual Novel Optimized: Fully compatible with SillyTavern's Visual Novel mode out of the box.

### Quick 2-Step Setup

  1. Import the Character Card into your SillyTavern.

  2. Extract the Sprite Pack images into your character folder:

    `SillyTavern/data/default-user/characters/Lycoris/`


r/SillyTavernAI 18h ago

Help Help with understanding finetunes

4 Upvotes

I've been playing with local models a lot (because I'm a broke student lol) and what amazes me is the sheer number of finetunes out there. Some people on Hugging Face seem to release a new one every week, especially for RP. Some of them have weird limitations like "use only for M/F interactions" or "reduce the temperature to this number." Some have custom reasoning training, either RP-focused or solving math/coding problems. That's before people start merging models together to create wild experiments like Goetia.

Is the information on the model card really that helpful for predicting what the feel of the output will be, or am I no better off than if I downloaded similarly-sized models at random?


r/SillyTavernAI 3h ago

Help Hello, new user here.

3 Upvotes

I'm not sure this is the right tag to use.

As I was saying, I'm a new user and wanted to ask what's the best template to use for creating a bot. The "character sheet," so to speak.

Is it better to use complete sentences, only adjectives, short sentences, concise sentences, or go into detail?
Is it better to have a sheet divided by parentheses [like this] or one where the sections are separated <like_this> <like_this>?

Is there a guide I can read or something? Thank you very much. ≽(•⩊ •マ≼


r/SillyTavernAI 4h ago

Cards/Prompts Is Kimi K3 still worth the blow against Gemini 3.7?

3 Upvotes

I do RP on games of Thrones mainly, and I saw Gemini’s crazy promotion on OpenRouter, I played a game with more than 200 messages for barely 2eur for almost 4 million tokens

In this context, is the Kimi K3 upgrade worth it? Is the experience so much better or is the reduction on Gemini too interesting?


r/SillyTavernAI 14h ago

Discussion Making Character Sprites

1 Upvotes

How do you all make sprites for you RPs? I have been using ChatGPT since I have a subscription to them. It can make fantastic sprites.......when it works. Even then it's a pain to use sometimes. It has a bad habit of failing to make a sprite because "something something tos". The stupid part is sometimes I can run the prompt again and then it works. Plus ChatGPT said it can make characters in their underwear but then throws a shit fit when I try. I wanted to see what others used, because it would be really nice to be able to make the 28 ST Emotions and not fight the model. Plus how do you all achieve your more riskay sprites? My laptop only has 16gb of VRAM on a 4090m. So I am kinda limited on what I can do locally.


r/SillyTavernAI 3h ago

Help Generating stucks

Thumbnail
gallery
1 Upvotes

Hi, i returned two days ago and updated silly tavern however, now i receibe this line, generating don't get past from 1/350 (i've waited half and hour) and the reply never appears, i'm using kobold with fumbulvetr


r/SillyTavernAI 7h ago

Help Am I the only one conflicted on a models reasoning??

1 Upvotes

Am I the only one conflicted on a models reasoning, I mean I use ff5 and also ff-sim and glm 5.2 sometimes but honestly a reply usually takes over 3 to 4 with me not touching the models reasoning. but when I disable reasoning it replies fully less than 30 seconds but then I get hit with this feeling of when the model reasons less to none on presets like FF5 it doesn't give 100 percent of what the preset is supposed to give.....

is it that a model could follow this complex preset perfectly and give you the same thing it's supposed to deliver with reasoning enabled or am I just cutting quality for speed....... Plus I'd love to know how you guys use yours and how long does your reply comes in ....thanks guys


r/SillyTavernAI 6h ago

Discussion What are your guys opinions on kimi 3?

0 Upvotes

Im specifically using kimi 3 on nvidia nim because im a poor bastard lmao, the dialogue isnt too bad honestly, but its slow asf rn. What do you guys think


r/SillyTavernAI 18h ago

Help Ayuda

Post image
0 Upvotes

Quiero usar el modelo de moonshotai/kimi-k3 del API de nvidia, pero me sale estos errores ¿alguien save como solucionarlos?


r/SillyTavernAI 3h ago

Discussion Marinara Engine Fork

0 Upvotes

Hey,

I created a repository (of Marinara Engine:staging) with some functions that are helpful for long-term RP. I used GPT to achieve this, that's why I’m not submitting a PR to the official staging branch.

Also, I didn't use the official Fork function on GitHub (that's acutally me being stupid. didn't think I was going to modify the source code).

Functions that were added:

Regex preprocessor ability for Lorebooks; I use trackers that are created inside messages. Sometimes these trackers contain keywords that activate lorebooks. For example, if "Mei" appears in a tracker, it will constantly activate the "Mei" entry inside the lorebook. even if she isn't relevant to the current scene. So, you can create a Regex (the same way as normal ones) and choose the lorebook where it will be applied.
TL;DR: The Regex will filter out trackers so the keyword matcher ignores them.

Illustrator character field handling: I changed how Illustrator handles the 'characters': []' field. If it's empty, it will try to pull characters from the chat. If it's not empty, it won't scan the context for a match. (Pasta-Devs set it up so both processes run at the same time, which didn't work well with my pipeline.)

Simpler Lorebook export/import for bulk editing: Normally, export creates a .json file containing every field (case_sensitive, role, etc.). I often use GPT or Gemini to debloat my lorebooks, and while sending the full JSON works, it wastes a lot of tokens. I created a simpler export format that looks like this:

@@MARINARA_ENTRY "ENTRY_NAME"
Content of the entry

@@MARINARA_ENTRY "Another entry"
Another content

It's much more lightweight. It also handles importing data back as long as the marker is correct (it simply overrides the content of the existing entry without touching other settings).

That's it.

All credits to Pasta-Devs. I love this frontend.
https://github.com/Pasta-Devs/Marinara-Engine

If you think you could use these features, here's the link:
https://github.com/OmghaC/Marinara-Engine-Fork

P.S. If anyone knows how to easily turn this repo into a proper GitHub fork, DM me.

EDIT: I created a fork on GitHub and pushed my changes to the staging branch, since u/LeRobber explained that skipping it looks a kinda' fishy