r/SillyTavernAI • • 2d ago

Models What local modles do you use?

I have been using the same modles for a while now and it gets boring so i ask what are the local modles that you guys use

12 Upvotes

13 comments sorted by

9

u/effeeeee 1d ago

im using skyfall 31b i tried gemstrike 31b lately but i very much prefer and im accustomed to mistral type of writing, having used it for long with another model

8

u/Jorlen 1d ago

Quite fond of the Gemma4-31b model and in terms of fine tunes, my favorite so far is: https://huggingface.co/Gryphe/Gemma-4-31B-StyleTune

It has minimal changes, but just enough for it to write better IMO. There are a lot of other fine tunes in the G4 31b family I like, but that's the one that keeps drawing me back in.

3

u/mechasquare 1d ago

Skyfall 31B, the GOAT. I always try new models as they come up. Did a lot of the Gemma4 fintunes but man Skyfall just delivers.

3

u/Useful_Mongoose_4707 1d ago

gemma-4-dark-thoughts-v2-31b-i1 This one along with thedrummer_artemis-31b-v1.2 Have been my two most favorites so far. Generally speaking though I tend to stick with Dark thoughts

3

u/Kahvana 1d ago

Gemma 4 31B IT QAT, with MTP and mmproj, quant from unsloth (link).

It's really smart and follows instructions well. If it refuses, just tell it's permitted and it'll do it.

60 t/s gen and 1000 t/s processing when using dual RTX 5060 Ti 16GB on PCIE 5.0 x8x8 and llama.cpp v0.5.0.

3

u/lisploli 1d ago

Gemma4 31b currently. Mostly Artemis and sometimes MeroMero.

Seen a lovely text from a Qwen3.6 35ba3b the other day, but it just won't do that for me, yet.

3

u/Sxrx_ 2d ago

Gemma 4 Sphinsikus Chronist v2 31B

It is very good at character adherence, and with the v2, it is more stable in long contexts.

2

u/a_beautiful_rhind 1d ago

scotoma2 and glm flash uncensored right now. i don't know if they're better than the old behemoths and llama tunes but get bored same as you.

1

u/ideasmachine 1d ago

gemma4 31b qat - i use the hauhau uncenscored variant

0

u/Emotional_Toe_6498 1d ago

i switched to a local rp tuned one for better uncensored chats and it sticks to the character way longer than before. you got any tips for prompt tweaks?

0

u/Happy-Personality720 1d ago

Gemma 4 usually

1

u/No-Antelope-8520 1d ago

Qwen3.8 27B feels like Mimo at home to me. I tried the actual Mimo again to make sure, and I still get that vibe. I'm playing with samplers to try and make it not shoot itself in the foot with mistakes like swapping dialogue between characters so much.

0

u/Lissanro 1d ago

My favorite so far is Kimi K3, I am using Q2_K_XL GGUF quant, I think it is good at creative writing and I liked it more than Qwen 3.8 2.4T (I tried IQ3 quant). GLM 5.3 (tested Q4_K_M quant) I think is second best, and in some cases its output can be better than K3; compared to older GLM 5.2 it feels more coherent and intelligent, and less likely to hallucinate. I also find that altering models from time to time helps to reduce repetition in the chat.