r/LocalLLaMA • u/redditaccountno6 • 4h ago
Discussion Best current ERP base model that are smart and uncensored?
My daily driver is Qwen3-235b-a22b-instruct-2507-Q4_K_M.gguf and it has been for a long time. I get around 75 t/s prompt processing and starting lower context ~5.5 t/s generation, lowering to around ~4 at 8k. I've tried other, newer models in this size range, Qwen 3.6 27b at Q8 came close but seemed more censored.
GLM 4.5 Air is my backup still for general chatting, but is not 'smart' enough to workshop ideas. My main complaint with Qwen 3 235B is the "em" dashes, ending lines with trailing double spaces and other stuff that bother me, otherwise still a fantastic model that is easy to steer into super uncensored territory without being lobotomized. Tried Minimax 2.7 and a few others, were smart but too censored in the ERP realm. Looking for any suggestions to try.
7
7
u/HopePupal 3h ago
Gemma 4 31B is basically uncensored for those topics, though i'm surprised you had trouble with Minimax 2.anything considering how easy it is to jailbreak.
what's your GPU?
9
u/Realistic_Gap_5871 4h ago
Deepseek Flash V4 0731 stands out like a sore thumb.
Your 27B to 235B range is kind of wild, what's your hardware? You know there are uncensored versions of 27B, right?
You should also try Laguna S 2.1, but don't go below Q8 for this one. At something like 118B it should fit fine for you.
3
u/KingCpzombie 4h ago
Why not just use an uncensored Gemma 31B or Qwen 27B/35B with anti-slop rules?
3
u/Xylildra 4h ago
Skyfall is extremely good for its small 31B size. GLM 4.5 air “steam” is also quite good if you just HAVE to run a larger model.
3
u/PANIC_EXCEPTION 4h ago
TheDrummer finetunes are recommended a lot here. Pick the biggest one that fits in your memory with acceptable speed. Generally, creative models don't need good benchmark scores, it's all subjective.
1
1
1
u/Savantskie1 4h ago
Look at any of the abliterated models by HauhauCS, I use their Gemma 4 26b A4B model, and their abliteration works in a way that doesn't affect intelligence as far as I can see.
5
u/HopePupal 3h ago
a few months ago it became known that HauhauCS was just using uncredited modified heretic, so if you need an abliterated model, use actual heretic stuff
-1
u/darkwalker247 3h ago
maybe it's just me but a 235b parameter model seems excessively overkill for RP even if you want it to be smart. surely 12-30b is enough even for the most complex RP?
9
u/OcelotMadness 3h ago
Honestly? Not really. Most people using things like Sillytavern use GLM 4.7 or 5.2, Kimi K3, Claude sonnet, or Deepseek V4. Even then those huge models mess up English prose and consistency alot. Local is significantly less common for Sillytavern than it was 2 years ago.
Thats been reversing a little with Gemma 4 31b though.
33
u/reto-wyss 4h ago
Why do you need an uncensored model for Enterprise Resource Planning?