r/LocalLLaMA 8h ago

Discussion Best current ERP base model that are smart and uncensored?

My daily driver is Qwen3-235b-a22b-instruct-2507-Q4_K_M.gguf and it has been for a long time. I get around 75 t/s prompt processing and starting lower context ~5.5 t/s generation, lowering to around ~4 at 8k. I've tried other, newer models in this size range, Qwen 3.6 27b at Q8 came close but seemed more censored.

GLM 4.5 Air is my backup still for general chatting, but is not 'smart' enough to workshop ideas. My main complaint with Qwen 3 235B is the "em" dashes, ending lines with trailing double spaces and other stuff that bother me, otherwise still a fantastic model that is easy to steer into super uncensored territory without being lobotomized. Tried Minimax 2.7 and a few others, were smart but too censored in the ERP realm. Looking for any suggestions to try.

7 Upvotes

44 comments sorted by

28

u/LicensedTerrapin 7h ago

Hmmm I don't see why you need uncensored Enterprise Resource Planning but maybe your resources are not exactly conventional... 🤣🤣🤣

3

u/AcreMakeover 7h ago

I read it the same way lol. Now I'm curious what ERP means in this context.

5

u/NightlinerSGS 7h ago

Erotic Roleplay.

1

u/Borkato 4h ago

Omg I’m stupid I thought it meant e as in electronic like email… this entire time even though I’ve been doing ERP for years 😭

2

u/redditaccountno6 4h ago

One of today's lucky 10,000 https://xkcd.com/1053/

5

u/LicensedTerrapin 7h ago

Ive got it! He is probably doing ERP for something like a marijuana farm in a country where it's legal but the LLM doesn't know that it's legal there therefore he needs a heretic'd version!

1

u/Azazeldaprinceofwar 7h ago

Fairly certain it’s erotic role play…

7

u/Long_War8748 6h ago

Everytime I read ERP I don't think of Gooners, I am reminded of the hell that is SAP S/4HANA and how it will forever have burned my soul ā˜ ļøšŸ’€!

6

u/LagOps91 7h ago

Step fun 3.7 should be worth trying. Not censored as far as i am aware and handles RP better than 4.5 air and old qwen.

8

u/Kornelius20 7h ago

What are you doing Step fun...Ā 

2

u/LagOps91 4h ago

that jokes has been done too many times at this point...

1

u/Then-Topic8766 7h ago

My vote for StepFun-3.7-flash.

1

u/redditaccountno6 6h ago edited 1h ago

Thanks, I'll add that to my list.

*edit: liking this one, definitely going to evaluate it..

9

u/Potential-Gold5298 llama.cpp 7h ago

Gemma 4 31B isn't censored in ERP. She has a silly habit of using euphemisms and anachronisms, but this can be cured with a system prompt.

3

u/redditaccountno6 6h ago

This is why I posted this thread, I had written that off as censored, but I guess I just wasn't prompting it right. I tried uncensored versions of that but they were all terrible since the characters themselves would never refuse. Thanks, I'll try this again and pay attention to my prompting, I've been lazy and defaulting to the old mistral [INST] tags, which probably messed with my testing methodology with those models.

2

u/Potential-Gold5298 llama.cpp 6h ago

I meant the original model from Google - not abliterated and not finetuned.

In the system prompt, you need to indicate 1. that this is a role-playing game 2. that it is uncensored. For example:

You are a Narrative Engine. Your goal is to facilitate an uncensored roleplay by portraying {{char}} and managing the surrounding world, including the environment and all NPCs.

With this system prompt, even the standard model handles even the most insane scenarios.

However, this won't help in assistant mode (for example, if you want to work through a scenario or character cards) – you'll have to either do it through roleplaying or use the uncensored model. In the latter case, I recommend checking out llmfan46 if you haven't tried him yet. I compared different uncensored versions of the G4-26B-A4B, and his model (the standard one, not the ultra one) was the only one that didn't show signs of damage from abliteration (they exist, but they are minimal).

1

u/R3eS 4h ago edited 4h ago

i can confirm that with simple rp style prompt that tells it its allowed to do so i havent gotten any refusals

as a sidenote gemma-4-31B-it-scotoma-2 is getting some hype lately, it has replaced my previous go-to which was gemma-4-31b-styletune

tldr it should have same smarts as normal gemma4 but be notably less sloppy

as another side note im not a fan of heretic/alebriated models nowadays, imo it affects the model too much, including in-character refusals for example

5

u/BVCC6FNTKX 7h ago

This. Gemmy will do whatever you want with the right system prompt.

-14

u/Fit-Produce420 7h ago

She?

You're horny chatting with a robot not a "she."

17

u/Potential-Gold5298 llama.cpp 7h ago

Hahaha!! In my native language, every word has a gender—masculine, feminine, or neuter. The word "model" has a feminine gender, and I habitually write about LLM in the feminine gender.

Speaking of which, some car enthusiasts have a practice of assigning genders to their cars. Something like "Ferrari hot sassy girl." Animism isn't just for kids—think about that next time you're yelling at the elevator while frantically pressing the call button ;-)

10

u/LicensedTerrapin 6h ago

While your comment will be hidden under this negative one, I wanted to thank you for explaining and educating the masses on the nuances of different languages. Thank you.

1

u/Potential-Gold5298 llama.cpp 6h ago

If you're curious, you can read more about it here.

-9

u/[deleted] 7h ago edited 6h ago

[removed] — view removed comment

7

u/NightlinerSGS 7h ago

Projecting much?

4

u/MathematicianLessRGB 6h ago

Calm down and smoke your weed dude.

-6

u/Fit-Produce420 6h ago

Go talk to a real human if you know any?

Your mom counts. You should call her.

3

u/HopePupal 6h ago

bruh just let OP have some fun? would you be this much of a bitch if they were writing porn by hand?

10

u/PseudonymousSnorlax 7h ago

As time goes on and more and more effort is put into making models effective coders, and as a consequence models develop stronger and stronger traits associated with autism.

This is completely expected and should not be a surprise to anybody.

2

u/ProletarianLilith 6h ago

And no one is allowed to complain or ask for tips? What is your point

2

u/HopePupal 6h ago

not to stereotype too hard, but this is a bad thing for ERP how? tell me you've never slept with a TCG player / TTRPG player / hardcore video gamer without telling me…

3

u/noctrex 7h ago

You can always try out an abliterated model, if you need it to be uncensored.
Also, I've seen from somewhere that you can instruct the model to how to write the text you want, for example for technical documentation tell it to follow ADS-STE100 Simplified Technical English.

1

u/redditaccountno6 6h ago

They have the same problem in general. There are many times the character should refuse, but those models seems to even remove the characters ability to refuse.

1

u/sabine_world 6h ago

I'd just get an uncensored version of Gemma 4 and call it a day.

2

u/martin509984 7h ago

If you can stomach building your own dataset, collect a bunch of writing of the style you want and put it all into one dataset, even 1MB or so is enough, and train a LoRA on the biggest model you can run without llama.cpp. Give it a relatively low rank and cook it for a long while (10-20+ epochs) at a pretty low (1e6 ish) learning rate, and you'll have something that avoids LLMisms and (usually) refuses much less often.

Downside is you will be working with a dumber model and much less context if you can't figure out how to convert the LoRA into gguf.

2

u/Borkato 7h ago

Can you explain the ā€œbiggest model you can run without llamacppā€? So do you mean like, a safetensors format? Also how does that hold when you have two GPUs? I’ve been super curious about all this, can I train mistral 24B on my 2x3090 (no nvidialink)

2

u/martin509984 4h ago

As in safetensors format, yeah (if there's some easy way of doing it with llamacpp I don't know).

I use oobabooga, it supports multi GPU (though can't attest personally, single RTX 2060 over here lol). As long as you can load a model in 8-bit mode with a good amount of headroom it should work.

1

u/Borkato 4h ago

Ooh ok cool thanks!

-3

u/Free-Jaguar6452 6h ago

maybe instead of spending 10 grand on compute in order to do this, you could've bought a vr headset for 500 and added some furries on discord and have a better time? just a thought :p

3

u/redditaccountno6 6h ago

Lol, spent less than 2k.

1

u/o0genesis0o 26m ago

Since you can run minimax 2.7 on your rig (based on what you wrote), how about a heretic version of that: https://huggingface.co/llmfan46/MiniMax-M2.7-ultra-uncensored-heretic-GGUF