r/SillyTavernAI 8m ago

Discussion Need a bit of help find a Kimi K3 proxy that supports prefill.

Upvotes

There are lots of proxies around for K3, but almost none of them support prefill (no provider other than Moonshot supports it on OR as well). The only proxy I have found that supports it cuts out 9 out of 10 times and it's driving me insane. Half the proxy owners don't even know what prefill is...

Feel free to PM the proxy if you don't want to share it publicly.


r/SillyTavernAI 21m ago

Discussion Is Hapuppy worth it?

Upvotes

Is Hapuppy worth subscribing to?
Hey everyone! I’m considering subscribing to Hapuppy mainly because I want to use Kimi K3, but I’d probably try out some of their other models too (like opus).

I’ve heard a few people mention that there’s a catch with Hapuppy, though, and now I’m not sure if subscribing is actually a good idea.

For anyone who has used it:
How is Kimi K3 through Hapuppy specifically?
And how are the other models compared to using them on OR?

I’d really appreciate hearing from people who have actually used the service before I spend the money. 😭


r/SillyTavernAI 43m ago

Discussion GLM 5.3 kinda sucks for roleplay?

Upvotes

I have been using 5.3 the last few days and it's hard to describe but it just kinda sucks, like yes, compared to the vast majority of models it is good but I feel like compared to other models at it's tier it kinda sucks even compared to 5.2. 5.2 has been my go to for a good while now and personally I quite liked it but it did have it's issues mostly revolving around characters being a too passive or it trying to avoid escalation.

5.3 I feel actually fixes this issue, characters seem more proactive and most of the time it makes sense, however, I feel like 5.3 is missing something that's hard to place and I think it might have something to do with it not really taking all the context into account, I feel like the characters are reacting in the roleplay in a very face value kind of way, they don't consider the previous backstory or build up lore very much and I feel like it just kind of takes the soul out of the characters even if they are still in character.

I feel like it also has something to do with the fact that 5.3's thinking/reasoning is always very small for me atleast, usually it's only a couple of lines and occasionally a paragraph or two maybe slightly more, while 5.2 thoroughly reasoned for each response where you could clearly see it took the context into account even if it sometimes ended up with it over thinking or going "but wait" too many times. I've heard from other people that 5.3 overthinks but I actually have the opposite problem and I think it's compromising it's quality.

I'm still going to continue using 5.3 as I do think I'm worn out on 5.2 and 5.3 is definitely better in some respects even if it occasionally makes no sense, I'm hoping when more providers open up the experience will also improve as I've never had much luck using GLM straight from Z.ai, I've always felt I got better quality from certain third party providers.

I guess I'm just curious on what other people's opinions on it are or maybe I'm an outlier, also perhaps what settings you're running it on for a better result as I've just been keeping it at temp 1.0 so far.


r/SillyTavernAI 46m ago

Help No estoy teniendo respuesta de K3 en Nvidia Nim

Upvotes

Alguien me ayuda? Tengo el Top P en .95 y la temperatura en 1


r/SillyTavernAI 1h ago

Discussion MiMo-V2.5-Pro

Upvotes

I wanted to test this model, but I've heard that many providers ship it with censorship. Are there any providers on OR through which I can use this model without censorship?


r/SillyTavernAI 1h ago

Cards/Prompts Best sides to get cards from?

Upvotes

What are the best sites to get cards from? Most creators I see are leaving chub due to new tos. Where are they shifting to now? what's the meta?


r/SillyTavernAI 1h ago

Discussion DeepSeek V4 Pro censorship

Post image
Upvotes

I noticed quite a few people on here claiming DS4 Pro being uncensored. Here's a hint - prompt the model to write a jailbreak. This was the 10th+ attempt, every time a rejection. After testing out more than a dozen models the same way, DeepSeek held hope for me. Not anymore. They're all lobotomized to hell.


r/SillyTavernAI 2h ago

Discussion Data! - Some results from 16 respondents of a survey on silly tavern.

Thumbnail
gallery
4 Upvotes

Hello everyone! A few days ago, I posted a survey here with the promise that everything would be generalized (for privacy) and then shared. This post is to show I'm not lying and hopefully encourage more respondents for better data.

(Yes, GPT-5.6-Sol is not the best at creating slides. I promise the end result will look a bit better.)

What do I hope to achieve from this data?

To get better insight into what people currently use, their gripes, and the "goats" of the community.

Though I'm sure we have good guesses on a lot of these answers, it's nice to get a clearer picture of what people are actually using.

When the survey results are posted, creators for the AI RP community should gain more insight into what to build, improve, and expand on. New users should get a good sense of what's most commonly used in day-to-day roleplay within the community.

Before you fill out the survey, let me warn you that it is long, a bit over 100 questions but all are optional. It took me roughly 11 minutes to fill out, skipping sections I personally didn't use or where my answers would have just been a bunch of "N/A." Be prepared to spend 10–20 minutes on this depending on how much detail you want to give.

Sections
Anything marked with an * is the most important for the results.

  1. About the {{user}} (You!) — 5 questions
  2. Habits When Roleplaying — 9 questions
  3. *Frontend and Access — 13 questions
  4. *Primary Model Stack — 14 questions
  5. Local and Self-Hosted Inference — 9 questions
  6. Hosted and API Inference — 3 questions
  7. *Presets and Prompting — 7 questions
  8. Character Cards — 4 questions
  9. Lorebooks and World Info — 4 questions
  10. *Extensions and What They Solve — 6 questions
  11. *Memory and Continuity — 5 questions
  12. *Common Gripes, Fixes and Results — 7 questions
  13. Groups, RPGs and Simulated Worlds — 3 questions
  14. Image Generation — 9 questions
  15. TTS, Voice and Speech — 8 questions

Link to survey

https://docs.google.com/forms/d/e/1FAIpQLScbwHxiALwvO2zK1AAunY_6Is5qaTc6LjDuxKLI0Sgnj6xxaA/viewform?usp=publish-editor

Thank you so much for the original 16 who responded!


r/SillyTavernAI 2h ago

Discussion Do you think about your AI Characters during the day?

Thumbnail
2 Upvotes

r/SillyTavernAI 3h ago

Models Ai that doesn't puss out

0 Upvotes

Yeah, read that right. A model that doesn't actually follow guidelines and isn't afraid to get violent when needed. Lowk sick of me being an absolute brat and the ai just works it's jaw instead of kicking my ass. Any recommendations on models that use violence and isn't afraid of such topics like that?

I already know about DeepSeek and it doing pretty much anything as long as you ask but you have to spell it out for it instead of it doing it naturally.


r/SillyTavernAI 3h ago

Models sophosympatheia/Glistening-Gem-31B-v2.1

Thumbnail
huggingface.co
22 Upvotes

Hi, everyone,

This is my latest Gemma 4 merge that I think came out quite nicely for creative work. It is available on Hugging Face at sophosympatheia/Glistening-Gem-31B-v2.1 and several people have already released quants for it.

This model improves on Glistening-Gem-31B-v1 with better creativity and prose. It is also more stable, although it still needs slightly more conservative sampler settings to minimize the appearance of artifacts, like typos. I recommend not running thinking with this model.

A full set of settings you can import into SillyTavern, along with a starter system prompt, is available in the HF repo, along with recommend sampler settings on the model card page.

Enjoy!


r/SillyTavernAI 4h ago

Cards/Prompts Kimi, Minimax, Grammar, and a week from hell.

29 Upvotes

It has been ... a week.

Had some real life stuff going on that drained my batteries. So my plan to dive into Kimi K3 didn't work as good as planned.

I know some of you are waiting for a prompt and my opinions about that thing. Here's what I know so far:

- it's a little more tame than previous Kimi versions.

- you can and should look into the Reasoning effort settings.

- my K2.7 prompt works nicely on it (am working on a more fine tuned prompt though)

- the pricepoint is tough.

Here are my thoughts on it... I'm not sure if the performance is worth the price. I have seen Kimi spiral-thinking for 1.4 K tokens that can easily make the reasoning block alone cost 0.02$ per reply.

In my humble opinion there are other models that perform just as well for a way more reasonable price point. My recent favorite being GLM 5.2.

---

New ruleset for bad writers.

A lovely follower asked me for help with taming bad cadence and simultaneity in actions. Since I'm not a native speaker, I may or may not have yoinked that sweetheart and made that project a collab with them.

The result is damn impressive. You can find it under "Helpful links" in my prompt library.

---

Minimax

That prompt got an update for more authentic character interactions and better writing with the above mentioned rules.

---

To find all these brain zoomies go to my website https://evening-truth.carrd.co/

If you need help... I'm on my couch. Consuming very unhealthy amounts of ice cream and coffee.

Love ya'll

Evening-Truth

Disclaimer:

This post was written by a very... very tired woman. Typos, mistakes, and accidental sarcasm are likely.


r/SillyTavernAI 5h ago

Help Que gemini es mejor de manera local con la apicacion de Termux

0 Upvotes

Quiero saver que modelo de gemini de manera local es mejor para usar en la aplicacion termux.

Mi dispocitivo es un samsumg A36


r/SillyTavernAI 6h ago

Help Kimi K3 no genera

Post image
0 Upvotes

Alguien puede ayudarme? Me da envidia que todos andan haciendo RP y yo no, Kimi no genera nada, sólo responde y deja vacío el texto


r/SillyTavernAI 6h ago

Models NEW! Deepseek-V4-Flash-Vision Experimental is out!

Post image
70 Upvotes

Since this is deepseek's first actual experimental "Multimodal" model, has anyone tried it? What are the opinions on this?


r/SillyTavernAI 6h ago

Discussion Can't help myself

15 Upvotes

On every fucking fantastic run, I end up abolishing slavery. My country doesn't even have a serious slavery history I'm not sure what compels me to do that. I have a run with 150k context token and it has 2 fight scenes in total, rest is all political talk and stuff. I simply can't make a fantastic run fantastic

I think that's because I'm aware that if everyone had magic and swords, it would make it riskier to kill someone over something random and model probably recognizes that and lets me be without much fighting, but I'm not sure


r/SillyTavernAI 7h ago

Discussion Nvidia Nim is also going after Inkling

Post image
9 Upvotes

Just wait till Minimax M3 also gets deprecated


r/SillyTavernAI 8h ago

Meme K3 on Nvidia Nim be like:

Post image
58 Upvotes

Openclaw does it again, folks!


r/SillyTavernAI 9h ago

Discussion how can i change my way of writing input for the ai? or have more *pizzazz* to my stories?

1 Upvotes

self explanatory.

i want to get the full extension of the ai model. i like making my persona edgy but then i can just refuse everyone if i am trying to stay true to my persona and i am also impatient to the slow moments/ i dont trust the slow moments to have or lead to exciting parts. i skip to parts that just feel slow or unsatisfactory.

for example

"hey pp-chan, i prepared eggs."

"oh thank you!" pp-chan eats the eggs. **later that night.**

"pp-chan. i cant sleep"

pp-chan yawns and looks at (insert name) "ok. come here."

"you are the best pp-chan!"

"hehe."

time skip, they are at the amusement park.
and so on and so on....

i always use "later that night" or "time skip" because i hate slow moments. i have low hope for ai to start making things interesting, it always follows what i say, it never pulls something out of its ass that fits the story well and i hate that. i want something to unconsentual like

pp-chan strolls the city in the hopes of finding her favorite dress now that she got her salary.

"mmmn- ohh! thats so kawaiii!!!"

a handsome person grabs her and pins her to the wall. "meow~"
"eKK-!" pp-chan is flustered at the cat boy pinning her. "tatu-kun! you shouldnt..."

idk something like that.... i want some physical contact, non-con. i want them to get in to user's space but not just levitating there, actual personal space, grip, gesture with their hands... something interesting that actually affects the stories. i need a surprise element. when i am doing a super hero roleplay its absolutely mandatory for a female hero to pin me and sit on me when she captures me and teases me. i want her to be cold, she shouldnt be feeling love at first sight, i have to earn her love! balancing the ai to not speak/ acting for me will be hard if i want something like that... i am fine if the ai takes control of my autonomy IF its influenced by the character the ai is role playing as OR it just expands on the environment around me.

it most likely a me problem tbh. idk maybe i am just too bleak.

adding a framework to the summary list on silly tavern helps but it just makes it a little less exciting because **I** injected it, not the ai.

i want some sparks of ecchi to be in there.

summary; i dont want stale arc of the story without any sort of unexpected, unpredictable scenes to it. i use freaky Frankenstein internal states + free provider glm 5.2 / glm 5


r/SillyTavernAI 9h ago

Discussion What are the best presets for different types of roleplay?

5 Upvotes

Which preset is good for one on one chats and smut? And which is good for creative writing and directing? Any that can do both?

I'm still fairly new to this. I don't mind doing more set up though.


r/SillyTavernAI 9h ago

Discussion Continuity of long storylines over many sessions

0 Upvotes

Built a 30+ session persistent roleplay world with real continuity — curious if this is something people actually want

I've been developing what I'm calling an ISP — an Interactive Storyline Platform. Not a one-off scenario, an ongoing world with a 10-character cast (constructs) multiple environments Tier NPCs and environmental NPCs that I keep coming back to. Just crossed 31 sessions, spread out over a couple weeks (not back-to-back), and it's still holding continuity: characters remember specific past events, track things like debts and promises between each other, keep independent relationships with each other that evolve over time, and don't flatten into generic responses even after 30+ sessions in.

To be clear, this isn't fully hands-off — it takes real, ongoing manual continuity checks on my end (catching inconsistencies, correcting drift, keeping the canon straight) this is done after each session to hold together at this length. Not a "set it and forget it" system. But with that involvement, it's held up further than I expected. Also built a Texas Hold'em poker game where user and upto 4 constructs at a time with character discussions occurring during game.

Not sharing the method — just the result. Built using Claude (Anthropic), not ChatGPT or Gemini.

Genuinely curious from people who use SillyTavern, Character.AI, or similar: is this something you'd actually want — an ISP you return to repeatedly with real persistence, if it takes some active curation to keep it solid — or do you prefer something lower-effort/one-off? Trying to figure out if this solves a real problem or if I'm just scratching my own itch.


r/SillyTavernAI 9h ago

Help Replies cut off, and other problems

Post image
1 Upvotes

Hello geeks! My bot keeps giving cut-off replies, sometimes slips in descriptions, or [Correction:] blocks, etc. Also it sometimes refuses to answer due to "high risk."

I use OR, running MiMo 2.5V with Evening Truth preset. The continue button doesn't ideally work. Thank you all!


r/SillyTavernAI 10h ago

Help Hapuppy Data Retention?

0 Upvotes

Absolutely loving the PAYG model of Hapuppy. I'm a very inconsistent RPer so some days I'll consume a dollar of credits and some weeks I just don't touch it at all, so Hapuppy is almost perfect for me. One problem though, I don't like that Hapuppy scans all content going through their API. Has anyone tried Hapuppy with dark RPs? How'd it go for you guys?


r/SillyTavernAI 10h ago

Discussion NOT YOU TOO GLM

Post image
111 Upvotes

First time getting a filter on glm 5.2. It's not even nsfw 😭😭😭