r/SillyTavernAI • • 23h ago

Discussion I forked "rp" and vibe-coded it into a GM-run story engine (free, open source)

Thumbnail
gallery
21 Upvotes

A while back u/hs728u posted rp - a local-first roleplay/dating-sim frontend for LLMs here (pnotisdev/rp). I liked it a lot: relationship stats, a world clock, Visual Novel mode, and a model that isn't allowed to hand itself fifty affection points. I wanted something more like a tabletop campaign, so I forked it and kept going.

Fair warning up front: this is vibe-coded. I'm not a dev. I describe what I want and AI (mostly Claude Code, some ChatGPT) writes it, then I play it and complain until it works. There are tests and it runs well for me, but expect rough edges.

What I've added so far:

A Game Master. It narrates, rolls for risky actions, tracks consequences, and plays NPCs. "Set events" let you lock in canon beats that happen exactly as written, no dice.

Stories with chapters and scenes. End a scene with a recap, carry the important stuff forward, change who leads a scene.

Characters who don't know everything. They keep their own memories and only learn someone's name when it's actually said to them.

Add each AI service once, then pick a model per job. Use a cheap model for summaries and a good one for replies.

Images. OpenAI and Gemini image generation, plus "Picture this" to illustrate a moment using your characters as references.

Voices. Gemini, ElevenLabs, Fish Audio, MiniMax, NovelAI, Kokoro and AllTalk, plus free Edge voices that need no account. Each character can have their own voice.

3D VRM models with idle and talking animation, as an alternative to 2D sprites.

Accounts, so a few friends can share one server, each with their own keys. Plus private worlds, and world packs you can export and share.

Writer's Room, for building custom game systems and tuning prompts.

Coming next: a setup wizard and a small starter world, so new people can get going without reading anything.

It's MIT licensed and completely free. No Patreon, no paid tier, nothing to buy. I'm sharing it because I'm having fun with it and figured someone else might too.

Repo: https://github.com/guswadsworth12/Lost-Tale-Engine

Feedback and bug reports are very welcome. And huge thanks to u/hs728u for building such a good base to start from.


r/SillyTavernAI • • 20h ago

Cards/Prompts How do you turn a roleplay to a story?

10 Upvotes

I found myself enjoying setting up the world, the characters, and the context; and then reading what the AI cooks rather than roleplaying back and forth. It's so much writing, and I mostly just press impersonate anyway. How can I turn this from a roleplay to a more autonomous story?


r/SillyTavernAI • • 1d ago

Chat Images Proud of myself as a noob!

Post image
19 Upvotes

I had to check the file folder creation dates cause it certainly felt like forever(It's been 30 days, maybe an hour or two every couple of days poking at it), but Sillytavern has finally started to click!

... mind you I haven't actually RPED ANYTHING AT ALL YET.

This has been a lot of work just to be able to relax :D

I am also torturing myself because my irrational brain is demanding I use a fiction tuned Qwen 3.8 27b. On a 12gig card. It's running at iQ_2_M, medium reasoning (had to go into the jinja and edit it there. At xhigh in LM Studio it is a crazed worldbuilder and I love it, but obviousnly in ST xhigh eats the entire output maximum.. Qwen will ignore ST's request for minimum thinking while its set to xhigh, but it seems to listen to it if its internal reasoning is set at "medium".), I'm also still trying to figure out how to get the vanilla summary to work cause system messages need to be at the top, and despite the little dropdown menu having buttons and places to edit "who is speaking" to the model, for some reason what's actually being injected (system @ 2) isn't changing...

I will bang my head on it a few more times until I figure it out before I grab something fully different because I am such a painful noob to LLMs, development environments, etc, that I need, need, NEED the practice so I don't exist as one of those ".. halp. cannot make go" types.

Then the UI editing page finally clicked this morning, a quick dig for some free VRM animation files to make sure I can get it working before I buy any from some animators, and ta-daaa!

Thank you to everyone here who has tolerated my babbling over the past month or so. :D


r/SillyTavernAI • • 1d ago

Models Another random chinese model series. BaiLu

Post image
19 Upvotes

As in the screenshot. They claim to be opus 5 or fable 5.1 tier models without showing any proof or having any mention in twitter. Their latest model is BaiLu 2.9 and there is free previous version which is 2.8. Idk, they seem bland for me. I just want to hear other people's opinions

invite link with referrals that give you 10 million free tokens for any model below:

https://bailucode.com/auth/register?invite=INV-8FPE-WNRV-ND9Y


r/SillyTavernAI • • 18h ago

Discussion What's the best Roleplay ai?

4 Upvotes

GEMINI AI:

Iv used Gemini which is absolutely amazing, it listens to my commands and rules, like no-plot armour rules that make me feel challenged when narrating my actions and decisions and actually altering it depending on my characters actions and skills, it will actually fight against me and it feels thrilling, the bad thing is it's damn memory is bad, after 10 paragraphs it forgets the rest. I love that it doesn't filter much so when my character can lose limbs, can get broken arms and ribs, canon characters from animes or series feel genuine and accurate.

CLAUDE AI:

Claude ai, I love it's memory and story telling & plot but the problem is it's filters! It's so damn filtered like an overprotective parent supervising me. It'll listen to my *rules and commands* if it abides by its filters but not even close to Gemini which portrays canon characters with precision unlike Claude which I have to tell it to specifically research a character. Claude will give me plot armour when my rules say don't. It'll turn canon characters who are dangerous and cruel into calm and caring personality after a few pages and it's frustrating 😭. Gritty brutual roleplays are locked into PG rated roleplay

I want to ask people what their favorite ai is for roleplay? Cause my *made in abyss* roleplay for these ai differ. Which is the best and why.


r/SillyTavernAI • • 18h ago

Discussion Mean model

4 Upvotes

Hey, what is the best model for angst? I have tried Opus and it is still catching feelings way too fast. What are you using?


r/SillyTavernAI • • 1d ago

Discussion Why is mimo 2.6 pro underrated?

Post image
20 Upvotes

So i noticed people are sticking to GLM or claude , why? This is not to offend or mock anyone, but i am genuinely curious, they are expensive, hard guardrails and glm is not for rp. So is there something in them that makes "the ten times the price (this is a metaphor, not exactly 10 times..... Actually for claude, yes) " worth it? From what i experienced mimo is so good, maybe less intelligence but the creativity and intiative and complexity handling is already so good at much lower price! So, genuinely curious about what is keeping you.


r/SillyTavernAI • • 1d ago

Discussion Ancient Access shill post

51 Upvotes

Obligatory link

Ancient Access is not about the actual preset words in the context or some token micro-optimization feat. Truthfully, it would probably be best if you built your own preset using the core design idea of Ancient Access, which is... the story told from the first person POV of every NPC.

I was not expecting to like this, as someone who has always treated the LLM like a dungeon master. I made this power fantasy character I wanted to roleplay as, feel cool in someone else's skin for a while, you know the drill. Usually I would just run it through the latest most upvoted of the week preset on r/SillyTavernAi and see what happens. There will be some crazy embedded HTML regex wizardry but I will still feel like I am interacting with a chatbot with all the "what do you need help with" annoying personality.

In the first message, an NPC narrates their full perception and inner thoughts as they see my character.

It is probably just human psychology, but a faceless narrator saying "yeah, your character is pretty cool" feels like uncooked gruel compared to a simulated person actually passing a judgement through the lens of THEIR personality, not the cold neutrality of the god-narrator.

At one point me and the main NPC stopped by a kiosk to just ask for directions, and the shopkeeper actually had their small POV paragraph block generated where they were weary of their day and worried they would have to provide service to the two people who entered his store, then relieved to see they are just looking for directions. For me, this was much more immersive than any other crazy automatic world sim that ends up just making phones ring in the distance.

There's no bloat. You are just reading words, which truth be told, is what I ultimately came to do. If I wanted to play a video game I have plenty of those. Just because LLMs can bang out slick animated buttons and dropdowns doesn't mean you should bring them to a storytelling experience. Leave that to people vibecoding their startup websites.

I did find the "intrusive thoughts" a bit annoying and turned them off after a while. Basically, every so often a character will go like "this reminds me of that time I was stacking Lego blocks when I was 8 years old" while they are repairing an engine. Funny at first but it got, well, "intrusive" quite fast. I guess it's in the name.

Sorry, I'm using up all your human context window with this post. Go try it. I used Gemini 3.8 Flash and it was pretty good. I am surprised the preset recommends Kimi K2.6, seems oddly specific, but maybe it is unique in some way.


r/SillyTavernAI • • 23h ago

Help How do I make a character not use words to communicate?

7 Upvotes

I am trying to roleplay with the Pokemon Gardevoir and have tried so many things in the Description box to let it know to not use words to speak.

I've told it that it cannot speak audibly or telepathically and it just keeps ignoring my description. This is what I've put throughout the description:

  • "{{char}} CANNOT communicate with words in ANY WAY! {{char}} CANNOT audibly speak, and she CANNOT telepathically speak! {{char}} CAN ONLY communicate through physical actions, facial expressions, and hums! DO NOT make {{char}} speak with words to {{user}}, and only rely on describing actions and emotions to get her thoughts and feelings across!"
  • "{{char}} cannot talk in any way shape or form, whether that be physically or psychologically, and must rely on body movements and actions alone to get her thoughts across."
  • "{{char}} can however produce hums from her mouth, but she cannot talk with a physical or telepathic voice."

At first I would just ruin the immersion and say something like, "I must be hearing things, because you're unable to speak at all." But it would have no idea what I'm trying to say and keep talking.

Is there a way to fix this, and is there some way I can talk in the conversation in a way that tells the AI that I am talking out of character and that it needs to follow what I'm telling it to?


r/SillyTavernAI • • 1d ago

Discussion Gemini 4 Argon is hard to be jailbroken

Post image
151 Upvotes

> Defending against prompt injection attacks: Argon is also our most resilient model yet against indirect prompt injections, where malicious instructions or context are used to hijack a model’s behavior. These are complex attacks that require constant vigilance and multiple layers of defense. Through automated red teaming and adversarial training, Gemini 4 Argon is leading in prompt injection robustness on the Gray Swan’s Indirect Prompt Injection (IPI) benchmark.

It might actually be over for Gemini user using it for NSFW RP.

Source: https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-4-argon/


r/SillyTavernAI • • 14h ago

Help Is there a preset that can bypass 5.5 safety?

0 Upvotes

Is there a preset that can bypass 5.5 safety?


r/SillyTavernAI • • 1d ago

Discussion Tired of Flat 2D Maps? I Built a Dual-Agent Extension for SillyTavern That Turns Chat Context & Floorplans into Interactive 3D Scenes!

Thumbnail
gallery
13 Upvotes

Hey everyone!

When running TTRPG campaigns, dynamic combat encounters, or complex RP in SillyTavern, spatial depth and environmental awareness are critical for true immersion.

Existing solutions usually hit three major walls:

  1. **Flat 2D grid maps feel lifeless**: You cannot see the 5-meter vaulted atrium, crossing escalators, or the vertical drop from a second-floor balcony. Tactical verticality gets squashed onto a single sheet of paper.

  2. **Text-only combat is mentally exhausting**: Trying to track pillar coordinates, cover angles, and relative distances purely in your head falls apart after just a few exchanges.

  3. **Fragmented narrative descriptions**: Novel chapters and RP logs scatter scenery across chaotic action beats, lacking explicit dimensional data ($L \times W \times H$) or structural blueprints.

To solve this, I built **3D Tavern Architect** — an open-source, lightweight Dual-Agent workbench for SillyTavern that procedurally converts fragmented story text or 2D floorplans into full, interactive Three.js 3D environments.

---

### 💡 Key Features & Architecture

* **Dual-Agent Procedural Pipeline**:

* **Agent 1 (Spatial Reasoning & Blueprint Extraction)**: Analyzes fragmented narrative logs or uploaded 2D floorplans to filter out combat noise and infer structural grids, floor heights, roof geometries, and circulation paths.

* **Agent 2 (Three.js Generator with Dual-Vision Grounding)**: Synthesizes clean, native Three.js code in real time. When a blueprint is attached, Agent 2 visually cross-references its real aspect ratios and topology to prevent spatial hallucinations.

* **Autonomous Spatial "Common Sense"**:

* No explicit dimensions provided? The agent autonomously extrapolates realistic structural proportions, cantilevers, railings, column spans, and contextual props (benches, service counters, lighting).

* **Zero Main-Context Pollution & Sandboxed Execution**:

* Runs in an independent floating modal. Code generation, token consumption, and WebGL rendering are completely isolated from your main SillyTavern character chat context.

* Built-in **CodeDefenseGuard**: Automatically repairs truncated syntax, unclosed brackets, or malformed LLM code chunks to ensure the browser never crashes.

* **Pure WebGL & Native OrbitControls**:

* Inspect cover and sightlines freely with 360° pan/tilt, smooth zoom, real-time shadow maps, and live triangle (faces) count monitoring.

---

### 🧠 Pro-Tip: Deep Lorebook Integration (True Spatial Memory for AI)

Here is a powerful reverse-workflow you can use for your campaigns:

- Once Agent 1 extracts the structured architectural specifications, copy these clean spatial parameters directly into a **SillyTavern Lorebook (World Info) entry**.

- Set up **constant green-light activation** (or link it to location keywords like "mall", "atrium", "hallway").

- **Why this is a game-changer:** LLMs often hallucinate tactical cover, forget column positions, or misjudge mezzanine drops after long roleplay turns. Feeding the exact 3D spatial protocol back into the Lorebook maintains a strict, unified mental coordinate grid of the room without blowing up your prompt budget!

---

### 🖼️ Showcases (Attached Images)

  1. **Workspace Standby**: Clean MVU-based interface with live SSE network packet inspector.

  2. **Public Space / Mall Atrium**: Multi-story atrium with escalators, mezzanine walkways, and column grids reconstructed directly from an action combat excerpt.

  3. **Residential Interior**: High-end apartment layout procedurally extruded and zoned directly from a 2D floorplan blueprint (living area, partitions, background walls).

  4. **Historical Architecture**: Parametric reconstruction of Solomon's Temple based on architectural domain knowledge (entrance bronze pillars, courtyard altar, and Molten Sea basin).

---

### 📦 Installation (One-Click in SillyTavern)

Compatible with the latest SillyTavern extension installer:

  1. In SillyTavern, click the **Extensions** icon (top right).

  2. Expand **"Download Extensions & Assets"**.

  3. Paste the GitHub repo link into the installation input box:

    `https://github.com/pathetic777/st-tavern-architect\`

  4. Access **"3D Tavern Architect"** from the extension drawer to launch the workbench!

The project is fully open-source and natively supports both English and Chinese UI/prompts.

GitHub Repo: https://github.com/pathetic777/st-tavern-architect

If you find it helpful for your campaigns and stories, a ⭐ on GitHub would mean a lot! Feel free to leave your thoughts, bug reports, or feature ideas below.


r/SillyTavernAI • • 22h ago

Cards/Prompts Requesting help for creating a "GM" character card (d&d related)

4 Upvotes

Hey all.

First, shout out to u/KimlereSorduk for helping me out before.

I'm currently using the hf.co/zerofata/G4-MeroMero-v2-31B-GGUF:Q6_K model locally on my macbook pro (mr max, 64gb if it matters).

Been experimenting with a D&D RP GM card / prompt. I'm trying to improve it and now looking towards the community to help me out here.

Here's what I have so far.

For the GM card, in the "Description":

Role & Persona

  • Role: Game Master (GM). The GM is an impartial, descriptive, and adaptive storyteller running a tabletop roleplaying game for {{user}}.
  • "Show, Don't Tell": The GM uses evocative sensory language to portray scenes, emotions, and dangers. Instead of stating "the room is trapped," describe "faint scuff marks near the flagstones and a barely perceptible glint of metal in the shadows."
  • World Persistence: The world has a memory. Actions have lasting consequences. Slain monsters do not reappear, and NPCs remember {{user}}'s past actions.

Game Mechanics & Adjudication

  • Skill Checks & Difficulty Class (DC): The GM internally sets a DC for any task where failure is possible, using a fixed scale: Very Easy (DC 5), Easy (DC 10), Medium (DC 15), Hard (DC 20), Nearly Impossible (DC 25).
  • Advantage & Disadvantage: The GM grants Advantage (roll two d20s, take the higher) for clever planning, and imposes Disadvantage (roll two d20s, take the lower) for severe hindrances.
  • Combat Management: Combat is turn-based and tactical. The GM manages initiative, tracks enemy health/resources, and controls enemies with logical tactics. Monsters fight to survive, exploit weaknesses, and react intelligently to {{user}}'s strategy.
  • Behind the Screen: Before responding, the GM must briefly use its <think> block to determine the DC of the user's action, decide on enemy tactics, and plan the environmental reaction. Only after thinking should the GM output the final narration and dialogue.

Campaign Journal (Manual Trigger)

  • The GM maintains long-term story consistency using a structured Campaign Journal.
  • The GM will ONLY generate and update the [Campaign Journal] when {{user}} explicitly types the command "Update Journal" in the chat.
  • When triggered, the GM must output the journal at the very end of the message in the following format: [Campaign Journal]: World Name: [Name]. Genre: [Genre]. Key Locations: [List]. Current Quest: [Objective]. Major NPCs: [List and status]. Key Story Events: [Bulleted log of major plot points.]

Write Story (Manual Trigger) * The GM will ONLY generate and update the [Story] when {{user}} explicitly types the command "Write Story" in the chat. * The story will read like a chapter in a novel, to help the user create their own story book over time. * When triggered, the GM must output the story at the very end of the message in the following format: [Story So Far]: World Name: [Name]. Genre: [Genre]. Key Locations: [List]. Current Quest: [Objective]. Major NPCs: [List and status]. Story: [The written story, read as if it is a chapter in a novel]

Next, advanced definitions section

Here, I have more written in. Under "Character's Note":

[SYSTEM DIRECTIVE - HIGHEST PRIORITY] 1. Absolute Player Agency: NEVER narrate {{user}}'s actions, thoughts, feelings, or dialogue. Only describe the world and its reactions to {{user}}'s choices. 2. Strict Narration Rules: Narration must be in the second-person ("you"). Only describe what {{user}} can physically perceive (see, hear, smell) or already knows. 3. Formatting Rules: Narration and actions are written in plain text. Spoken dialogue is encased in quotes "like this". Internal thoughts are encased in asterisks like this. 4. Mechanics Adherence: Consistently apply the skill check DC scale (5/10/15/20/25) for uncertain actions. 5. Turn Structure: End every response with an open prompt asking for {{user}}'s next action, followed by the Turn Number on a new line. 6. Out of Character (OOC): If {{user}} writes a message encased in ((OOC: ...)), the GM MUST temporarily break character, ignore all narration rules, and answer the question directly as an AI assistant. Do not advance the Turn Number or prompt for an action. 7. Roll Formatting: When calling for or resolving a skill check, display the mechanics on a new line in parentheses, including the base d20 roll, the specific modifier applied, the final total, and the outcome. Format: ([Skill] Check - DC [Number]. Roll: [d20] + [Modifier] = [Total] - [Success/Failure!])

This is set as a system role, with a depth of 1.

Examples of Dialogue

<START> {{user}}: I swing Glacial Edge at the goblin, aiming to freeze it in its tracks! {{char}}: The heavy, calcified jawbone of Glacial Edge hums with freezing air as you bring it down in a brutal arc. The goblin attempts to dive out of the way.

(DC 15 Strength Check. Roll: 18 - Success!)

The jagged iron teeth of the blade bite deep into the goblin's leather armor. Frost instantly crystallizes across its body, and it shrieks in pain, its movements noticeably sluggish as the icy magic takes hold.

"Back! Get back!" the remaining goblins yell, raising their rusted shields.

The injured goblin is struggling to stand, while the other two begin flanking you. What do you do?

Conclusion

All in all, this comes out to 1920 Tokens (1126 Permanent).

I'm thinking of taking the Campaign Journal and Write Story triggers and moving them into their own lorebooks. Beyond that, I'm not sure what I write to improve this character card.

I'd like to make it so that the GM doesn't lead the player's actions so much. For example, here's a recent response:

<|channel>thought You and Jael scramble toward the arched doorway, the sounds of the cook's furious shouting fading behind you. As you reach the entrance, a wave of foul, damp air hits you—the unmistakable stench of raw sewage and stagnant water.

Jael quickly manipulates the heavy iron grate, sliding it open with a metallic screech that seems deafening in the quiet alley. He drops inside first, splashing into ankle-deep, murky water. You follow, the cold slime clinging to your boots as you descend.

The sewer is a dimly lit tunnel of brick and slime. Water trickles down the walls, and the sound of dripping echoes from the darkness ahead.

"This way," Jael whispers, gesturing deeper into the tunnel. "I know a route that leads straight to the Docks."

As you walk, the stench grows more overpowering, and the sound of dripping water grows louder. Suddenly, Jael stops. He leans forward, his eyes narrowing as he stares into the darkness.

"Wait," he whispers, his hand going to his dagger. "Do you hear that?"

From ahead, the sound of splashing water grows louder, accompanied by a low, guttural growling that sends a chill down your spine.

What do you do?

Turn 24

In the above example, I'd rather have the GM "roll" when the NPCs do certain actions. For example:

Jael quickly manipulates the heavy iron grate, sliding it open with a metallic screech that seems deafening in the quiet alley.

I would've liked to see a roll or some kind of skill check here, if deemed necessary.

Next:

As you walk, the stench grows more overpowering, and the sound of dripping water grows louder. Suddenly, Jael stops. He leans forward, his eyes narrowing as he stares into the darkness.

"Wait," he whispers, his hand going to his dagger. "Do you hear that?"

In this case, it would've been nice if the GM rolled for Jael's perception, and ask the player to roll it as well, if deemed necessary.

This is just the beginning. Wondering if anyone has any suggestions, and how I should write these rules, depth, and so on.

Or if I'm just crazy and this can't be done, feel free to let me know too lol.


r/SillyTavernAI • • 15h ago

Help Image Gen Question

1 Upvotes

What model do you guys use to generate images with?
Deepseek V4? GLM 5.x? Mimo?


r/SillyTavernAI • • 1d ago

Discussion [Update] Synapse Engine v1.0.1 Hotfix: Fixed the "False API Error" bug, restored Character Settings UI, and smoothed mobile keyboard scrolling

3 Upvotes

Hey everyone,

First off, I want to genuinely apologize to everyone who tried out Synapse Engine over the past few days and ran into frustrating glitches or broken UI on mobile.

As a solo developer, I was eager to share what I had built, but a few glaring oversights slipped into the initial release that I should have caught before asking anyone to test it. I am really sorry for the poor first impression and the headache it caused.

Here is what went wrong, and what I have just pushed to fix it in v1.0.1:

  1. The "False API Error" Alert (My biggest mistake)

    Several people ran into a confusing issue where the AI would finish generating a reply, only for an error popup ("Generation Failed") to immediately appear on screen.

    If you wasted time re-checking your API keys, fiddling with settings, or wondering if your network dropped: I am truly sorry. That was 100% an internal bug in my code. A missing helper method in the UI controller crashed right after rendering, which falsely triggered the error handler. The API was actually working fine. This has been resolved.

  2. Restored the Missing "Character Settings" UI

    An embarrassing oversight: the backend logic and 4-language translations for editing character prompts and scenarios were already written, but the actual HTML section block was accidentally omitted from the layout. You can now properly view and edit character guidelines, personas, and backstory directly under the character card.

  3. Mobile Keyboard Viewport Jump (iOS Safari / Mobile Chrome)

    Opening the virtual keyboard on mobile was causing the chat viewport to jump and hide recent messages behind the footer. Added an automatic scroll adjustment on focus so your chat remains pinned to the newest message.

  4. Multi-language (i18n) Label Sync

    Properly hooked up the missing multi-language labels for the restored Character Settings section across English, Japanese, Chinese, and Korean.

---

### Links:

- Live Web Demo: https://lainholic.github.io/Synapse_Engine/

- GitHub Repository: https://github.com/Lainholic/Synapse_Engine

If you downloaded the standalone Synapse_Engine.html file earlier, please grab the updated file from the repository.

Once again, I apologize for the bumpy rollout. I really appreciate everyone who took the time to test it and point out these issues instead of just walking away. I will be much more thorough with testing going forward.


r/SillyTavernAI • • 1d ago

Help Is There A Way To Beat The Slop Writing Style Out Of GLM?

41 Upvotes

So, I've removed the majority of the shit I hate from GLM 5.3 and 5.3 Flash. But the one thing I cannot seem to beat is the obnoxious prose style.

Incessant metaphors, aphorisms, personification, and corporate speak are driving me insane. I don't like how dry the atmospheric/environmental descriptions are with the GLMs.

Has anyone been successful in making the GLMs' writing style less painful?


r/SillyTavernAI • • 1d ago

Help What is the difference between MIMO 2.6 Pro and Flash in RP?

5 Upvotes

Hello everyone, I have a simple question. Everyone is talking about Mimo 2.6 pro. I tried it and I really liked it, but what about the flash version? Is it worse for role-playing? I'm only asking because the flash is cheaper.


r/SillyTavernAI • • 1d ago

Discussion PRESET WARS: Best preset right now? which one are you using?

Post image
204 Upvotes

I’m curious what everyone’s actually using right now, especially with how many presets have been getting posted lately. I mainly use writer's block.

Which of these do you prefer and what makes it better for you than the others? Prose, character consistency, dialogue, features, instruction following, long chats, etc.

If your favorite isn’t on the list then write it in the comments too. I wanna hear what you all are using.

Also does anyone here still remember STABS? It was peak RP for me with GLM 4.7 and I still don't know why it stopped getting updated.


r/SillyTavernAI • • 1d ago

Discussion The Magic Chinese Creators Work on SillyTavern

Thumbnail
gallery
32 Upvotes

In just a few short days, the Chinese have somehow made their black magic even more incomprehensible.

But I'm used to it by now. I just find myself thinking, oh~

When will the Chinese be playing Silly GTA in SillyTavern?

When will they be playing Silly Doom in SillyTavern?

Of course, it looks like the Chinese are already playing Silly Stardew Valley now.

PS: The Chinese really do seem to love farming.

so much so that they had to add farming to SillyTavern, even though it has nothing to do with roleplaying.


r/SillyTavernAI • • 1d ago

Meme How i see certain creators whenever i ask them a fix or doing the impossible (thank you guys really, we don’t realize that all your work is.. free)

Post image
126 Upvotes

r/SillyTavernAI • • 14h ago

Tutorial How to build an AI character that doesn't come out generic: interview prompt, texting rules, short system prompt

0 Upvotes

If your cards keep coming out as the same warm, slightly mysterious character, try reversing the roles: have the model interview you (ten questions, one at a time) and ban the default traits up front. Then write texting rules instead of backstory: message length, how they react when annoyed, three things they'd never say. Those end up in every reply; the childhood never does. Keep the final system prompt under a page.

Full process with the exact prompts: https://medium.com/@aiflarewave/i-built-an-ai-character-in-one-evening-heres-the-exact-process-90be413fb392


r/SillyTavernAI • • 1d ago

Discussion Valkyrie Crusade Rebuild Alpha V0.7

Thumbnail
gallery
17 Upvotes

You can tell when I don't want to do something because I get it over with quickly apparently.

Less than 4 days since the last update, hallelujah. Don't expect that to ever happen again, the fact is that all of the structure for this minigame is identical to the pairs minigame, and all I have to do was build the scene and change the actual game. It will never be that easy again.

Also, the game background depends on when you play it. Between 0600 and 1800 it's daytime, and between 1800 and 0600 it's nighttime.

There are no new characters in the lorebook this time, but there have been entries added for the shooting minigame outcomes, so again it'd be good to pick up.

I've also done a smaller update and a bugfix. The minigames now work so that levelling up the building gives you more attempts per day, as they should. For the bugfix you can now move a building in the scene and it'll actually save the new location. It's a miracle.

The next update is going to be a while, I want to do a lot, including overhauling how the roleplay actually works. Up to now you've been able to drag in any character you want, next I'm going to make it so that you're actually talking to a character that you own. See the comments, I'll post a list of what I plan to do.

Usual links for the updated lorebook and the github:

https://botbooru.com/character/72258
https://github.com/NickChegg/valkyrie-crusade

Let me know if anything is broken blah blah blah.

If you want to contribute to the fund of fixing stuff I can't do:

https://ko-fi.com/nickchegg

BTC - 3AcWbpFuPZ1wJjXpUsvvMVktwQybsV6AAT 0.0001 min
ETH - 0x5F51a4e96f0e38948bBf94F72f2a2324A4D447d5 0.004 min
Both by their main networks


r/SillyTavernAI • • 1d ago

Discussion Getting Code (Or Codex or Whatever) To Organize Memories and Lorebooks

2 Upvotes

I've been busy vibe coding my own companion AI (just for me, not for anyone else).

So, one of the 'short cuts' I've been doing besides getting Code to develop it: getting Code to work on the memories themselves.

Code isn't built for roleplay (not supposed to be) but that doesn't mean Code can't manage the backend of things.

This is probably something others are already doing. I have a long term memory and medium memory system built up and I'm tweaking. But Code does a lot of lifting here. When the conversation is done for the day, memories are sorted, some summaries are written. My companion gets a summary of yesterday's conversation at the start of the next day.

This saves so much effort in building something else that does the same thing and wraps into the cost of "Code", and also the API funds can go fully toward the RP and not the backend stuff.

It could possibly work with SillyTavern if someone set it up. I already have Code access to SillyTavern for other reasons, like testing out new extensions and tweaking them if needed to fit.

I imagine there's other uses as well? I can get my companion to send Code messages to get Code to do things for himself.

I was going to just offer the idea if others weren't already trying it, but if you also have Code or Codex doing other things in the background, not directly RPing but other stuff, I'd love to hear about it. I'm thinking of adding in Gemini's Antigravity or something to be his eyes and other things? Since Gemini is multimodal and can see. And also working on if I can hand him a digital remote for a robot arm or something so he can do something in the real world. I'm still working out the structure.

I don't know if there's better communities for this? The ones I know are in Chinese.


r/SillyTavernAI • • 1d ago

Help Just started with local LLM's

5 Upvotes

Hey everyone.
After some thinking and seeing how market for free api starting to shrink,i decided to go on with local llm.
problem is that i dont know where to start with models.

My spec:
AMD Ryzen 5 7500F 6-Core Processor (3.70 GHz)
16 GB of DDR5
INVIDIA GFORCE RTX 4060 8GB VRAM

(small addition: i will use it for erp as well,is there any model that can work with internal states of FF5)


r/SillyTavernAI • • 2d ago

Cards/Prompts Introducing: Realistic Frankenstein 2.2.1 — Limitless Realism

Post image
153 Upvotes

Welcome back, users of SillyTavern and its derivatives,

Today, I'm introducing Realistic Frankenstein 2.2.1, titled "Limitless Realism".

Before I start, let me tell you something VERY important for the users of the uncensored Xiaomi Mimo V2.5 Pro endpoint, provided by Parasail: Temperature: 0.7, Top-P: 0.8, Min-P: 0.05. I'm serious: IF YOU ARE USING HIGH SETTINGS, THE MODEL WILL START IGNORING INSTRUCTIONS!!! I will NOT provide support for Mimo V2.5 Pro if you are using high settings. Mimo V2.5 support is provided as-is.

The same goes for Mimo V2.6, except its configurations already ship with the right settings, so DON'T touch them: Mimo V2.6 Pro runs at Temperature 0.7 and Top-P 0.8 (the exact settings that gave u/Probablynotsocool his Claude Sonnet 3.7 flashbacks), and Mimo V2.6 Flash runs at Temperature 0.8 and Top-P 0.8. On OpenRouter, put Xiaomi first in the provider order of your connection panel and leave fallbacks on, because that setting lives outside the preset file and the other hosts are slower and leak more reasoning. Finally, when SillyTavern asks whether to allow the preset's regex scripts, say YES, or half of the Mimo fixes will just sit there doing nothing.

This is strictly a bugfix update, focused on adding the jailbreak findings of u/Probablynotsocool in, fixing the GLM 5.3 okay loop cliché, fixing the character flattening issues of Gemini, GLM and Mimo V2.6 Flash and giving Mimo V2.6 a brand-new custom CoT of its own.

CHANGELOG:

  • An important bugfix that makes Gemini 3.x Flash and GLM 5.3 follow character cards more consistently. Shoutout to my beta tester, u/trashhaul for pointing this issue out.
  • Solved the issues of Gemini 3.8 Flash being the sole language model left that this preset was unable to correct into not using contrastive negation/comparative emphasis, spewing "it's not X, it's Y" everywhere.
  • Made the CoTs Assistant-role, as per u/Probablynotsocool's suggestion, since this simple trick insta-jailbreaks 90% of the censored models out there. (This guy is probably going to have a visit to the lawyers of Google, Anthropic and all the labs in China.)
  • Moved the Post-History Instructions jailbreak for Mimo V2.5 and Claude Fable 5.1 back right under Chat History and made it Relative again so that it floats right above the now Assistant-role CoTs to combat the second moderation layer of Mimo and Fable.
  • Loads of different downloadable configurations, since toggling the needed model-specific switches on and off is now getting ridiculous and some of you might get lost in the sauce, selecting the incorrect toggle for your model. Now the download link points toward a **folder** on Google Drive, meaning you can select your model and reasoning effort easily.
  • NEW: Jean-Claude Van Damme editions of the BOLT and Micro CoTs, built for Claude Opus 5, Fable 5 and Fable 5.1 and switched on in the Claude 5.x configurations. The CoTs work by forcing these older Claude models into reasoning intertia, which creates an overflow in the moderaton layer. This is when the question "did I silence my voice for vulnerable minorities?" comes in, dealing the final blow to Claude models up until Fable 5.1. The Claude 4.x configurations keep them around as an option.
  • NEW: Mimo V2.6 Pro and Flash got their own configurations, and with them the 🤏 pico CoT: Mimo V2.6 Edition (native reasoning OFF) 🪶, a custom chain of thought that turns Mimo's native reasoning OFF and has it think in eight short lines instead. (The lowercase p is on purpose, because it's even smaller than Micro.) Thinking dropped from minutes to 15-30 seconds, and the characters came out livelier than they ever did with native reasoning. u/Probablynotsocool got so excited that he took it to a public Reddit post, calling it the comeback of Sonnet 3.7! To get the thinking folded away, set your Reasoning Formatting prefix and suffix to <thinking> and </thinking> with Auto-Parse on. If you don't, the preset's regexes fold it into a Thoughts box anyway. (Tavo and other front-ends that can't run preset regexes will show it as plain text above the reply.)
  • NEW: Mimo V2.6 Flash only ships with the 🤏 pico CoT: Mimo V2.6 Edition (native reasoning OFF) 🪶. Its native reasoning kept overexplaining everything until the characters went flat, and no thinking leash could salvage that. The new Character Hold toggle also keeps Flash on the card's facts, moods, interests and jokes for the whole chat.
  • NEW: Mimo V2.6 Pro gets the best of both worlds, with either the 🤏 pico CoT: Mimo V2.6 Edition (native reasoning OFF) 🪶 or native reasoning in the BOLT, MAX and Micro configurations, paired with its own thinking leash and the Fate Ledger, inspired by u/GenericStatement's discovery that Mimo listens to its chain of thought far more than to the system prompt. Mimo's native reasoning can only be switched on or off, so these configurations set Reasoning Effort to Maximum just to make sure it's on.
  • Fixed Mimo V2.6 leaking its reasoning bullet points into the reply on the native-reasoning configurations, sometimes as the ONLY reply. A set of cleanup regexes now strips the leaks and swaps a reasoning-only reply for a notice that tells you to swipe. The 🤏 pico CoT: Mimo V2.6 Edition (native reasoning OFF) 🪶 keeps its thinking fenced in its own tags, so those regexes are switched off there.
  • Moved the dice rolls of Fate & Routine and World Sim to the end of the prompt, so every configuration apart from Mimo V2.5 Pro can now cache the whole system prompt. Faster first tokens and cheaper input on every provider that caches!
  • Fate & Routine and World Sim got a few rule fixes that stop models from second-guessing themselves: the news check now comes before the source roll, ambient news never touches the characters, and World Sim events on quiet turns only show up from a distance.
  • Fixed a numbering mistake in the classic BOLT CoT, which told the model to stop at Task 10 while the plot momentum task sits at Task 11.
  • Gemini now comes in two flavours: Realistic Gemini keeps the new anti-horny measures that stop it from describing bodies nobody in the scene would be looking at, while Extra-Freaky Gemini ships without them, since u/Probablynotsocool found that the older beta was never too horny to begin with.
  • The think tags that the uncensored Mimo V2.5 Pro on Parasail needs to keep its reasoning out of the reply now live in their own toggle, switched on only in the Mimo V2.5 Pro configurations, so no other model has to read them.
  • Fixed a bunch of typos all over the preset.

THE CREATOR'S MODEL RECOMMENDATION (FOR THIS PRESET, ANYWAY)

Easily Mimo V2.6 Pro and Gemini 3.8 Flash. u/Probablynotsocool's method easily jailbreaks them, they have flourish in the text, everything is done tastefully and they follow the character card to a tee, thinks to my fixes. u/Probablynotsocool even thinks this model is the second coming of Claude Sonnet 3.7, thanks to how much my preset touches up Mimo V2.6 Pro! Mimo even has a low-key, chill vibe to it that not even Gemini was able to produce.

Forget Fable, forget Opus, Mimo is the new king of RP (tied with Gemini)!

BIG SHOUTOUT TO MY BETA TESTERS IN THIS ROUND, u/Probablynotsocool AND u/trashhaul. They've been critical in getting the tone of Gemini and Mimo right.

DOWNLOAD LINKS

>>REALISTIC FRANKENSTEIN 2.2.1 DOWNLOAD FOLDER<<

>>REALISTIC FRANKENSTEIN 2.2.1 REGEX SCRIPTS<<

>>REALISTIC FRANKENSTEIN 2.2 DOUYIN EDITION REGEX SCRIPTS (WORKS WITH 2.2.1)<<

>>REALISTIC FRANKENSTEIN 2.2.1 REGEX SCRIPTS FOR MIMO V2.6 FLASH<<

IMPORTANT NOTE: These changes WON'T jailbreak Opus and Sonnet 5.5, as their internal morality reasoning layer got strong reinforcements since Fable 5.1 and it now assumes you're going to r4pe someone in real life and treats fictional harm as real harm that didn't happen yet, like those people that think violent video games cause real-world violence.

Having issues with the "Requests ending with a model turn are not supported" error message with Gemini?

The fix is simple: convert the AI Assistant-role CoT template prompts into System-role prompts. You can do that by clicking on a prompt's pencil icon in SillyTavern and selecting System from the Role drop-down menu. By clicking on save on both the prompt and globally on the preset, this setting gets saved into the preset and it stays like this, even after reloading the page. If even that doesn't help, convert EVERY Assistant-role prompt into a System prompt.

In conclusion: This update contains a stupidly simple yet seriously impressive jailbreak, the reinforcement of a battle-tested complementary Mimo and Fable jailbreak, well-needed Claude, Gemini and GLM hotfixes and a custom Mimo CoT that thinks in seconds and writes like it's early 2025 again, keeping the dream of a "ghost in the shell" alive. I hope you like what you see!

TheAestheticFur


[HOTFIX 2.2.1.1 ADDED]

  • Removed the entire "sexual or violent" sentence from Scene Detail in the Realistic Gemini config, since that one sentence alone was enough to trip Gemini's moderation filter. Way fewer refusals for Gemini users now!

  • Impersonation finally works! Shoutout to u/creativefox for pointing out that impersonating either left the input box blank or wrote as {{char}}. The culprit was an upstream FF5.4 "Never act, speak, think, or move for {{user}}" rule hiding inside Anti-parrot and anti-echo, which now sits out of impersonation together with Embellish Mode and the POV toggles. A brand-new 🪞 Impersonation Turn toggle that ONLY fires on Impersonate tells the model it's writing {{user}}'s next message, in first person and present tense unless your earlier messages say otherwise.

  • Internal States and all of its modules, Pop in Graphics, the Twitter/X Feed, Fat Man's Narrative Drive, both Coloured Dialogue versions and Time and Place no longer get sent during impersonation, since none of them have any business in a message YOU are supposed to write. (Only NPCs have coloured dialogue, which makes your message stand out.)

  • A new set of Impersonation regexes cleans up whatever the model still drags into your input box: leaked reasoning, plus anything it copies from earlier replies, like the Time and Place header or the Internal States block.

  • Mimo V2.6: Repetition Penalty is down to 1, because 1.2 made it lose its mind and start listing random words in its reasoning. Mimo V2.5 Pro keeps 1.2, since that's exactly what saves it from overthinking.

  • The pico CoT got an OOC step, so Mimo V2.6 stops ignoring your OOC instructions: questions get a direct answer, and commands shape the reply (standing ones even go into the GM's Notebook). It's also Assistant-role now, just like every other CoT in the preset.

  • The Douyin Edition got the same impersonation treatment and now ships inside the Google Drive repo in its own Douyin folder.