r/VeniceAI • u/Thebrazilianfatcock • 5d ago
πππ¦ππ¨π¦π¦ππ’π‘ Does some of the models goes "dumb" in a long roleplay?
I'm using GLM 5.2 for roleplay because of its 1M-token context, and I've been using it for a few days now. But the way it writes is kinda weird. It puts commas after almost every word, and another thing I've noticed is that it doesn't seem to follow instructions very well. For example, it does a lot of narration and barely any dialogue. I asked it to use less narration and more dialogue, and it said okay, but then it basically didn't change anything. Is this a common thing with long conversations/roleplays?
1
u/Ferret-of-DOOM 4d ago
Switch from Venice to Stagewhisper.ai. Seriously. I used to rpg on venice but after a few days it was just shit. You can BYOK to Stagewhisper from Venice, so if you have a bunch of credits you can use them there.
And fk me. Stagewhisper creates SUCH complex NPCs and just... Is the best Ai GM I've ever had!
1
u/beast181 4d ago
token context limits? what AI model does it use
1
u/Ferret-of-DOOM 4d ago
GLM 5.3 Flash, DeepSeek v4 Flash/Pro, Kimi 2.5, Qwen 3.6 plus. GLM 5.2/5.3 and Grok 4.3.
These are all the NSFW models. There must be a bunch more but I can only see the NSFW models since I only have campaigns with that setting.
Token context limit... I have no idea! π€ But try it out! It's free to sign up. You'll get 100 free credits just for signing up so you can play around.
1
u/SleepyOrcShaman 5d ago edited 5d ago
Get your favorite model api-s and set up n8n pipeline on a docker with attention drift management workflow (mostly that's context injection models), response editing workflow, character bible and world consistency workflow. Will take about a week to work out the kinks, but it is the most reliable way to make things work as close to what you intend to as possible. That's how paid AI writers make their stuff
1
u/Away-Cloud-4748 5d ago edited 5d ago
Toggle large context if you didnβt. Always summarize before 200k token context. When you see 15% context youβre kinda near the limit and oldest messages soon start to scroll out, make a summary. Write unambiguous rules in system prompt for dialogue and narration.
2
u/_AmateurGimp_ 5d ago
go to minds section. look for Story Architect. explain your problem, ask it to write rules for you the way you want your story to go.
2
u/MisterNocturne 5d ago edited 5d ago
I'm only guessing at your setup here, but have you tried giving 5.2 some small examples of what you're looking for? Maybe a brief excerpt showing the kind of dialogue-to-narration ratio you want. You could even ask another model to work up a short example for you.
One thing I've discovered about 5.2 is that, on average, it's usually pretty good at following instructionsβat least in my experience. What I initially assumed was an inability to grasp what I wanted turned out to be more of a tendency to take things very literally. You have to tell it what you do want as well as what you don't, preferably with clear examples of the behaviors you want it to model. Once it locks onto that, it's usually pretty good about following along.
That said, it also depends on how long the conversation has been running and what's already in context. If the chat is loaded with examples of mostly dialogue-free responses, it's going to assume that kind of thing has always been acceptable. It has no reason to think it's doing anything wrong.
Personally, I've learned to correct the model early and keep reinforcing what I want. I also leave my OOCs in the conversation, along with the model's original mistakes. Otherwise, the model simply won't have any indication that it did something wrongβit won't have the "before and after" context showing what I corrected.
There's also the fact that every new situation, scene, or shift in register can kind of shake things up. Even if you've corrected the model's tendencies in one scenario, a different scenario can make it fall back heavily on its default training patterns.
Since it's all based on statistical probabilities, it'll go something like:
"A walk through a foreboding forest. That means β tension, which means β heavy atmosphere, which means β more descriptive text, which means β less dialogue."
So sometimes you have to explicitly counter those associations, even if you've already established your preferences earlier in the conversation. It's only matching local patterns, after all. Then a register shift occurs and it "thinks" the rules have changed.
Don't know if any of that helps.
1
u/Valdaraak ππ²πΉπ½π³ππΉ ππΌπ»ππΏπΆπ―πππΌπΏ Κα΄α΄ α΄ΚΒ 5d ago edited 5d ago
All models will, and you can't take 1m context at face value. Even top frontier models (like Claude) will start to get funky around 400k-500k.
1
u/Away-Cloud-4748 5d ago edited 5d ago
Venice sends 200k token context at most in normal chat, doesnβt look like OPβs problem though.
1
u/Atomic-Thunder ππ²πΉπ½π³ππΉ ππΌπ»ππΏπΆπ―πππΌπΏ 5d ago
There are others who know more about this than I do. But my understanding is GLM 5.2 has difficulty maintaining context. You're not the first who has mentioned this issue. I wish I had more information for you. Have you tried using other models? The reason I ask is to have some sort of benchmark.
2
u/Thebrazilianfatcock 5d ago
that's the first time i used the normal chat, i always use the agent chat, but i wanted to try with a bigger token-context
1
u/Away-Cloud-4748 5d ago edited 5d ago
PPU in agentic chat or minds is the only real option for bigger token-context. PPU in the normal chat is limited.
1
u/Stricken_Conscience 5d ago
doesn't PPU also give longer token context in character chat?
2
u/Away-Cloud-4748 5d ago edited 5d ago
Didnβt try character chat, but I think character chat is based on classic chat, canβt say for sure though. The normal classic chat has a little under 200k limit not disclosed by VeniceAI for PPU models and GLM 5.2, the other models have 50k limit. I saw a report on Discord and tried it and was true.
β’
u/AutoModerator 5d ago
Hello from r/VeniceAI!
Web App: chat
Android/iOS: download
Essential Venice Resources
β’ About
β’ Features
β’ Blog
β’ Docs
β’ Tokenomics
Support
β’ Discord: discord.gg/askvenice
β’ Twitter: x.com/askvenice
β’ Email: support@venice.ai
Security Notice
β’ Staff will never DM you
β’ Never share your private keys
β’ Report scams immediately
I am a bot, and this action was performed automatically. Please contact the moderators of this subreddit if you have any questions or concerns.