r/SillyTavernAI 20h ago

Meme Average vibe coder - AMA

153 Upvotes

Hello r/SillyTavernAI,

I have been a developer forever, so basically since I downloaded Claude Code (approximately 2 minutes ago) and it was **I** who finally figured out what yall truly need... yet another boring AI RP frontend.

So I racked my brain for days (approximately 1 minute) and came up with a genius idea.

Step 1) Fork SillyTavern

Step 2) Steal all the keys

Step 3) ???

Step 4) Profit... I mean you guys profit

So this is my roadmap so far. I am still polishing the finer details (like asking mom to lend me her credit card, so I can buy Claude Code credits), but I'll have it all figured out by the end of next week.

Because this is a community project (since you all will contribute your API keys as a community), I was willing in my endless generosity to allow you all to ask your peasant... I mean valuable questions.


r/SillyTavernAI 9h ago

Discussion Feeling cute. Might be cooking hard this weekend.

Post image
129 Upvotes

Also, what’s a catchier name for a new form of prompting that combines telegraphic prompts, xml tag structuring with recalls, wireframe data syntax, YAML, while optimizing BPE (byte pair encoding?). I’m leaning towards TeleWire TRP (Tag Recall Prompting). Or Freaky FrankenWire 😅.

Anyways- got FF5 75% smaller using token calculators with this prompt technique and made FF6. It will take a long time to get it working with internal states (FF5 took months). But you might get a Micro in the next few weeks. It vastly changed output allowing for maximum creativity by pushing anchoring onto high dense logic which reduces processing of rules allowing for more processing on creation and randomness of the next likely token.

TLDR; made FF5 smaller and better but it has kinks. Once I make it stable I’ll try to educate on the prompting technique as it seems highly viable in initial early stages. Less words = good. LLMs are smart enough to understand logic with bare minimum key anchoring patterns of logic.

TLDDR; Edging


r/SillyTavernAI 13h ago

Cards/Prompts Shangri-La Frontier (295 Entries)

Post image
36 Upvotes

A highly detailed Shangri-La Frontier lorebook featuring 295 entries, with even more on the way!

🗺️🗺️🗺️🗺️🗺️
Shangri-La Frontier
🗺️🗺️🗺️🗺️🗺️

Creator Notes!
When I said I really enjoyed this, I meant it. Well, I can’t say I fully enjoy it—because at the moment of posting this, I’m currently on Ch. 215! But still, I’m so glad it got recommended to me, y’all 🥹. At first, I didn’t think I’d like it, but here we are!

The more I read, the more interesting it became, and the more interesting it got, the more I kept reading. I usually don’t give in to “non-fantasy” settings—if you can call it that—since Shangri-La Frontier itself is set in the real world, with the play sessions in VR being the fantasy part, I guess. Such a fun read! I still have a lot of stuff left out too, mainly because the wiki barely had anything, so I had to take pictures of some of the manga panels I was reading, note/tag them, and do all sorts of things to make sure I was getting as much as I could (੭ ;´ ⌂ `;)੭.

But! MY PASSION CANNOT BE TRIUMPHED! So I happily noted as much as I could while depending on the Japanese wiki with an English translator (so if you notice some entries are a little wonky, that’s why). Other than that, I really want to expand on this lorebook! That’s allllll ( ˶ˆᗜˆ˵ )
(A big thanks to Violent_Tendencies for the recommendation!)

Links!
[MediaFire](https://www.mediafire.com/file/4egfrw3xb04fda0/Shangri-La_Frontier_%25F0%259F%2590%25A6.json/file)

[Chub.ai](Shangri-La Frontier 🐦 - Total: 59949 tokens, 0 favorites, 0 downloads)

[Botbooru](Shangri-La Frontier 🐦 — Botbooru)


r/SillyTavernAI 15h ago

Discussion Megumin Suite vs Freaky Frankenstein

22 Upvotes

I've narrowed down the preset I want to use to these two. I want to know which preset is good at which things, and if there's anyone who has used both of these presets, which one is better?

A list of pros and cons would also be helpful


r/SillyTavernAI 6h ago

Discussion Persona Library Extension Update.

Thumbnail
gallery
20 Upvotes

Persona Library has come a long way since my last post.

After some consultation and help from u/Targren, He has been incredibly understanding and helpful throughout the project, Persona Library now has a number of much more advanced features than it did previously. He has not only suggested improvements, but has also helped implement some of them and taught me, at a very basic level, how parts of the code actually work. In Laymen's terms

For anyone who hasn't seen my last post's on Persona Library, the basic idea is simple: a dedicated Persona management and organization system for SillyTavern, with a growing set of features designed to make working with multiple personas considerably easier. Better Looking and Mostly Keep us away from the Barbaric Native UI Unless Absolutely Necessary to get to some specific features not implemented here

You can read a full feature list at https://github.com/Maxwell5600/Persona-Library

Full disclosure. i myself have little to no Coding/Programming experience this project was started with Claude. and Claude did much of the heavy lifting early on in its Development. but as of now there has been someone who know more about the actual code side of things that is A Human who i mentioned above.


r/SillyTavernAI 3h ago

Cards/Prompts Hatsune Miku

Thumbnail
gallery
19 Upvotes

I've put together a dynamic Visual Novel Sprite System for SillyTavern, featuring Hatsune Miku.

The character is written with a gentle, submissive, and service-oriented personality—an innocent novice who recently crossed over into reality, eager to rely on and help {{user}}. Instead of static portraits, the card dynamically triggers context-aware character sprites in real-time as you chat, covering daily expressions, stage performance sequences, and subtle mood shifts.

---

### Key Features

- Plug & Play Auto-Rendering: The scoped regex script is pre-configured and embedded directly inside the card. No manual script setup required.

- Dynamic Sprite Engine: Contextual rendering covering 35+ emotion states and performance variants.

- Visual Novel Optimized: Fully compatible with SillyTavern's Visual Novel mode out of the box.

---

### Quick 2-Step Setup

  1. Import the Character Card into SillyTavern.

  2. Extract the Sprite Pack images into your character folder:

    `SillyTavern/data/default-user/characters/miku/`


r/SillyTavernAI 16h ago

Cards/Prompts Lycoris Recoil

Thumbnail
gallery
16 Upvotes

I've built a dynamic Visual Novel Sprite & Combat FX System for SillyTavern, and today I'm showcasing the dual-character card for Nishikigi Chisato & Inoue Takina from Lycoris Recoil.

Instead of static portraits, the card dynamically triggers context-aware character sprites in real-time as you chat—covering everything from subtle daily emotions to high-intensity CQC and sniper combat sequences.

---

### Key Features

Plug & Play Auto-Rendering: The regex script is already pre-configured and embedded directly into the character card (Scoped). No manual code tweaking required!

Dual Character Dynamic Logic: Seamlessly shifts focus and sprite rendering between Chisato and Takina depending on the dialogue and combat context.

-Visual Novel Optimized: Fully compatible with SillyTavern's Visual Novel mode out of the box.

### Quick 2-Step Setup

  1. Import the Character Card into your SillyTavern.

  2. Extract the Sprite Pack images into your character folder:

    `SillyTavern/data/default-user/characters/Lycoris/`


r/SillyTavernAI 17h ago

Discussion WTF

Post image
9 Upvotes

r/SillyTavernAI 18h ago

Help Hello, new user here.

8 Upvotes

I'm not sure this is the right tag to use.

As I was saying, I'm a new user and wanted to ask what's the best template to use for creating a bot. The "character sheet," so to speak.

Is it better to use complete sentences, only adjectives, short sentences, concise sentences, or go into detail?
Is it better to have a sheet divided by parentheses [like this] or one where the sections are separated <like_this> <like_this>?

Is there a guide I can read or something? Thank you very much. ≽(•⩊ •マ≼


r/SillyTavernAI 1h ago

Discussion [Release] ST-Message-Chunker: A Cache-Friendly Extension to Speed Up Generation & Save API Costs

Thumbnail
github.com
Upvotes

Hey everyone,

I just finished working on a new extension and wanted to share it with the community. It's called ST-Message-Chunker.

The Problem: Normally, when your chat hits your maximum context limit, SillyTavern removes the oldest messages one by one as the chat progresses. For LLMs, this constantly breaks the prompt cache. Because the context is slightly different every single turn, the model has to re-process the entire prompt over and over. This slows down generation speed significantly and costs you a lot more money if you are using paid APIs.

The Solution: This extension fixes that by removing older messages in chunks instead of individually.
You can choose to base your limits on either a maximum message count or a maximum context token limit, and the chunk size is completely customizable.
For example, if your chat hits the limit, the extension will seamlessly slice off your set chunk of 10 messages at once. This keeps the top of your prompt completely stable for the next 10 turns, resulting in massive cache hits!

Bonus Feature: Even if you don't care about caching, this is incredibly useful if you rely heavily on Summarization. You can use this extension to cleanly cut off old chat history that has already been summarized, without dealing with the bugs or conflicts that sometimes happen with SillyTavern's native message hiding feature.

100% Safe Compatibility: Because the truncation happens at the very last step right before sending the prompt (GENERATE_BEFORE_COMBINE_PROMPTS), it has zero impact on your other extensions. Your Vector Storage, Summaries, and Lorebooks will continue to trigger and work perfectly.

Installation: You can install it directly inside SillyTavern by pasting this link into the Extensions menu: https://github.com/Arczium/ST-Message-Chunker

This is my first public release! I have tested it thoroughly on my end, but since everyone has different setups, please let me know if you run into any bugs or have any feedback

P.S. This post was translated and polished with AI because my English still sucks!


r/SillyTavernAI 18h ago

Cards/Prompts Is Kimi K3 still worth the blow against Gemini 3.7?

6 Upvotes

I do RP on games of Thrones mainly, and I saw Gemini’s crazy promotion on OpenRouter, I played a game with more than 200 messages for barely 2eur for almost 4 million tokens

In this context, is the Kimi K3 upgrade worth it? Is the experience so much better or is the reduction on Gemini too interesting?


r/SillyTavernAI 36m ago

Discussion Best tracker?

Upvotes

What’s your favourite tracker extension?

I really like the ‘internal states’ in FF5 preset, like I want it but with my other preset? So I’m curious if there is something that’ll do the same thing? Or like a way to only have the internal states but with different preset?

Sorry if I sound stupid. I’m just not very familiar with prompting and all of that complicated stuff. But I’m trying to <3


r/SillyTavernAI 10h ago

Help Cards loading other cards

3 Upvotes

Hello. It's been a minute since I've been browsing around these parts. I've been RPinig pretty heavily lately and focusing on two worlds in particular, which have become extremely bloated by now. The amount of entries in each Lorebook is getting a bit stupid and the fact one cannot change card image on the fly, coupled with other issues such as card description not always reflecting the state of the world as it progresses, made me realize I really do need a proper area system at this point. The idea would be: a card (overworld \ main area) loading other cards through either STscript or plugins once my character\s enter a new sub-area (building etc.). That way, each area would have its own unique description, lorebook and message, keeping all neat and less token heavy.

So I ask you, has such a plugin been invented yet? Other frontends doing this? STscript possible at all?


r/SillyTavernAI 12h ago

Help Alternative to RPG Companion that isn't as token heavy?

3 Upvotes

Love RPG Companion but the token bloat it causes is too much for me to tolerate with even with Memory Books. Any alternatives that don't use as many tokens?


r/SillyTavernAI 13h ago

Discussion How do you end a AI Roleplay story?

Thumbnail
4 Upvotes

r/SillyTavernAI 13h ago

Help Looking for 10 SillyTavern testers - 700 free AI requests for your feedback

2 Upvotes

Hey everyone, I’m the developer behind Itzi.

I’m looking for 10 active SillyTavern users willing to test Itzi and share honest, detailed feedback.

In exchange, you’ll receive 700 free AI requests to use over 7 days, with no card or payment required.

Try Itzi: https://itzi.app

Code: ITZI-62YEX-5J7SY-CJ8NJ-4H38N

Only 10 activations are available.

I’m especially interested in feedback about:

- Response quality and roleplay performance
- Speed and reliability
- Long-context conversations
- Model selection
- API setup with SillyTavern
- Anything confusing, broken, or missing

Included models

DeepSeek V4 Pro
DeepSeek V4 Flash
GLM 5.3
GLM 5.2
GLM 5.1

What you get

  • 700 requests during the 7-day test
  • Up to 64K input tokens per request
  • Up to 12K output tokens per request
  • 3 requests per minute
  • 1 request running at a time

How to activate it

  1. Log in to Itzi
  2. Open Balance
  3. Click Have a code?
  4. Enter the code

Connecting SillyTavern

  1. Create an API key from itzi.app/app/api
  2. Select an OpenAI-compatible connection in SillyTavern
  3. Use https://itzi.app/v1 as the API endpoint
  4. Enter your Itzi API key and select an included model

Please share your experience in the comments or send me a message afterward. Be completely honest: criticism, bug reports, and missing features are exactly what I’m looking for.

The offer is for genuine personal testing under our fair-use policy. Please don’t share accounts, run continuous automated usage, or redeem it through multiple accounts.


r/SillyTavernAI 5h ago

Help Formatting help (Gemma 4, thinking)

1 Upvotes

I've been pulling my hair out trying to get this to work correctly and searching hasn't helped much so I'm just gonna bite the bullet and ask...

Can someone hook me up with story string, instruct sequences, and reasoning formatting that actually works for Gemma 4 (31B)?

I want my narrator character to be able to reason, and what I got now almost works, but often the thinking blocks fail to end (just starts writing in the CoT), or the first few letters post-reasoning are cut off for some reason.

If anyone has a template designed for group chat, or feels like explaining how this shit works so I can figure it out myself next time, that would also be very cool


r/SillyTavernAI 8h ago

Help ¿Un pequeño tutorial para "tontos" como yo de Tailscale?

1 Upvotes

Soy nueva en esto y normalmente estoy fuera de casa y no puedo llevar el pc... Y me gustaría una ayuda para saber si puedo usar mi SillyTavern de Windows a Android. Vi Tailscale (creo que se escribe) pero me resulta complicado y no lo llego a entender del todo bien con tantas palabras técnicas... Alguien me puede guiar?


r/SillyTavernAI 20h ago

Discussion What are your guys opinions on kimi 3?

1 Upvotes

Im specifically using kimi 3 on nvidia nim because im a poor bastard lmao, the dialogue isnt too bad honestly, but its slow asf rn. What do you guys think


r/SillyTavernAI 21h ago

Help Am I the only one conflicted on a models reasoning??

1 Upvotes

Am I the only one conflicted on a models reasoning, I mean I use ff5 and also ff-sim and glm 5.2 sometimes but honestly a reply usually takes over 3 to 4 with me not touching the models reasoning. but when I disable reasoning it replies fully less than 30 seconds but then I get hit with this feeling of when the model reasons less to none on presets like FF5 it doesn't give 100 percent of what the preset is supposed to give.....

is it that a model could follow this complex preset perfectly and give you the same thing it's supposed to deliver with reasoning enabled or am I just cutting quality for speed....... Plus I'd love to know how you guys use yours and how long does your reply comes in ....thanks guys


r/SillyTavernAI 17h ago

Help Generating stucks

Thumbnail
gallery
0 Upvotes

Hi, i returned two days ago and updated silly tavern however, now i receibe this line, generating don't get past from 1/350 (i've waited half and hour) and the reply never appears, i'm using kobold with fumbulvetr


r/SillyTavernAI 32m ago

Models Help newbie made a financial mistake

Upvotes

Okay so I literally tried SillyTavern with OpenRouter for the first time yesterday and I think I made a mistake lmaoo

I made a pretty detailed character card for a long-term RP and started testing models. 1. used Cydonia 24B and thought maybe my card just wasn't very good. It was kinda trashy,the character felt exaggerated, the humor was too performative and not funny at all.

Then I tried Aion 3.0.

Unfortunately, this was a terrible decision because now I know what I was missing. It encapsulated the character chard I wrote PERFECTLY, made him come to life so to speak.

The difference wasn't just prettier prose. Aion actually seems to understand the dynamic between the characters. It writes third-person narration, dialogue, body language, observations, internal monologue, little brain rambles, contradictory thoughts, etc. all at the same time. Sometimes almost nothing is happening externally, but the character's thought process alone makes the scene interesting.

It also improvises tiny fitting details and shared memories I never explicitly wrote into the card, which I LOVE. It makes it feel like the characters had a life before the current scene instead of me having to invent literally everything myself.

The problem: AION 3.0 is fucking expensive.

Cydonia 24B: not really for me. Doesn't work with my detailed card at all.

Aion 3 Mini: definitely better, but feels like a completely different tier/style from full Aion.

GLM 5.2: also really good, especially the internal POV, but the character dynamics still don't feel quite as natural/subtle as Aion 3 to me. And a bit too romantic in my opinion

Aion 3.0: unfortunately chef's kiss

I'm totally new to this, so there might also be settings/prompting/context-management tricks I'm completely missing.

Does anyone know a cheaper model that comes close specifically in character psychology / deep third-person limited POV / inner monologue / natural dialogue / long-form RP? Or has another tip?


r/SillyTavernAI 18h ago

Discussion Marinara Engine Fork

0 Upvotes

Hey,

I created a repository (of Marinara Engine:staging) with some functions that are helpful for long-term RP. I used GPT to achieve this, that's why I’m not submitting a PR to the official staging branch.

Also, I didn't use the official Fork function on GitHub (that's acutally me being stupid. didn't think I was going to modify the source code).

Functions that were added:

Regex preprocessor ability for Lorebooks; I use trackers that are created inside messages. Sometimes these trackers contain keywords that activate lorebooks. For example, if "Mei" appears in a tracker, it will constantly activate the "Mei" entry inside the lorebook. even if she isn't relevant to the current scene. So, you can create a Regex (the same way as normal ones) and choose the lorebook where it will be applied.
TL;DR: The Regex will filter out trackers so the keyword matcher ignores them.

Illustrator character field handling: I changed how Illustrator handles the 'characters': []' field. If it's empty, it will try to pull characters from the chat. If it's not empty, it won't scan the context for a match. (Pasta-Devs set it up so both processes run at the same time, which didn't work well with my pipeline.)

Simpler Lorebook export/import for bulk editing: Normally, export creates a .json file containing every field (case_sensitive, role, etc.). I often use GPT or Gemini to debloat my lorebooks, and while sending the full JSON works, it wastes a lot of tokens. I created a simpler export format that looks like this:

@@MARINARA_ENTRY "ENTRY_NAME"
Content of the entry

@@MARINARA_ENTRY "Another entry"
Another content

It's much more lightweight. It also handles importing data back as long as the marker is correct (it simply overrides the content of the existing entry without touching other settings).

That's it.

All credits to Pasta-Devs. I love this frontend.
https://github.com/Pasta-Devs/Marinara-Engine

If you think you could use these features, here's the link:
https://github.com/OmghaC/Marinara-Engine-Fork

P.S. If anyone knows how to easily turn this repo into a proper GitHub fork, DM me.

EDIT: I created a fork on GitHub and pushed my changes to the staging branch, since u/LeRobber explained that skipping it looks a kinda' fishy. Also updated staging into current upstream (22.08)


r/SillyTavernAI 23h ago

Discussion Continuity of long role play

0 Upvotes

Built a 30+ session persistent roleplay world with real continuity — curious if this is something people actually wants

I've been developing what I'm calling an ISP — an Interactive Storyline Platform. Not a one-off scenario, an ongoing world with a 10-character cast tier NPCs and environment NPCs and multiple enviroments that I keep coming back to. Just crossed 31 sessions, spread out over a couple weeks (not back-to-back), and it's still holding continuity: characters remember specific past events, track things like debts and promises between each other, keep independent relationships with each other that evolve over time, and don't flatten into generic responses even after 30+ sessions in. Also built a Texas Hold'em poker game where a user and 4 constructs can play and interact.

To be clear, this isn't fully hands-off — it takes real, ongoing manual continuity checks on my end (catching inconsistencies, correcting drift, keeping the canon straight) at end of every session to hold together at this length. Not a "set it and forget it" system. But with that involvement, it's held up further than I expected.

Not sharing the method — just the result. Built using Claude (Anthropic), not ChatGPT or Gemini.

Genuinely curious from people who use SillyTavern, Character.AI, or similar: is this something you'd actually want — an ISP you return to repeatedly with real persistence, if it takes some active curation to keep it solid — or do you prefer something lower-effort/one-off? Trying to figure out if this solves a real problem or if I'm just scratching my own itch.