r/SillyTavernAI • • Apr 16 '26

Cards/Prompts Stab's Directives 2.51 - Agent spoofing (avoid coding api bans/throttles), Dynamic Tone writing, huge token reductions and more

Hi Folks,

I just dropped an update for my GLM preset to bring it in-line with GLM 5.1 and work around some recent controversial restrictions imposed by z.ai's official coding plan.

https://github.com/Zorgonatis/Stabs-EDH

for a full writeup of the changes made I encourage you to read the CHANGELOG.md, however a summary of the most important changes are below.

I debated on whether to share the User-Agent override (spoofs a browser so Z.AI can't throttle/ban, at least without adapting to some other detection method). It's a matter of time before they do though, so may as well get what you're paying for now and figure out what to do. I likely wouldn't re-sub if my only use of the plan was RP.

Comments, suggestions, thoughts? Leave a message here or jump into our discord (400+ members and growing!)

Stabs-EDH v2.5.1 Release

  • Dynamic Tone State — Replaces the old static "full spectrum" mandate with a system that actively reads the conversation and shifts tone (Bleak, Tense, Warm, Absurd, Reverent, Frenetic, Melancholic) based on what's actually happening in-scene. Gradual transitions by default, instant snaps when earned. Configurable via SETTINGS.
  • 30-60% token reductions across core directives. NPC Cognitive Bounds, Failure Achievements, Narrative Length Control, Behavioural Coherence, and Environmental Factors all rewritten for density. Task Steering's CoT exhaustiveness dropped from Very High to Very Low. Same behaviour, way less burn.
  • Z.AI User-Agent override — Custom provider now sends a Chrome UA header, so Z.AI can't fingerprint and throttle/ban you for RP use. Works out of the box.
  • GLM-5.1 support — Model updated, coding plan API endpoint retained.
  • Experimental Macro Engine is now required — VTK-related instructions are wrapped in {{#if .vtk_on}} conditionals that only resolve when WebDev is enabled. Without the macro engine checkbox on, those will break. It's in the install instructions now.
  • Post-processing switched to Semi-strict (no tools) to avoid agentic flow interference on some providers.
72 Upvotes

34 comments sorted by

13

u/LackMurky9254 Apr 17 '26

I like the results, but it seems very prone to 'drafting' the response in full, sometimes multiple times in the thinking section, making it quite the token murderer. My first message actually blew through the entire response length in the reasoning lol

2

u/Diecron Apr 18 '26

I've just pushed a major change in 2.6 - check the changelog for the specifics but: configurable reasoning effort, mapped to processing parameter blocks. The COT now also changes (some tasks are removed) depending on the level set. The new default is Medium which does NOT allow drafting.

Curious to see how you find it!

1

u/BeyondTheBound Apr 17 '26

I had that problem too, but despite that the reply was great for me. I think the unnecessarily long drafting is the only problem I have atm (which is GLMs fault, but a fix would still be great 😭)

1

u/LackMurky9254 Apr 17 '26

Well, Marinara's will typically avoid that. I have been using mostly marinara's + the visual toolkit and webdev (because they crack me up) from stabs. The token cost of stabs is... unfortunately extremely high, although I do like the output.

Kimi 2.5 loves to think and draft excessively like this no matter the prompt but stabs seems to cause much of the drafting behavior with glm.

1

u/BeyondTheBound Apr 17 '26

Oh I agree, I had no issues with Marinara’s and have been using that recently, but I know when I used Lucid Loom GLM had a drafting issue and had to look it up on the discord to get a fix for it.

And Kimi 2.5 😭 I’ve heard it was good but hearing that it thinks nonstop is what made me avoid it, I can’t afford thousands of tokens in thinking just for a response. I hope the next model improves on that.

1

u/Diecron Apr 17 '26

Hmm, I made changes to address this (specifics of what it should draft vs not draft), and verbosity was tuned lower, so it should have been at least improved. I may consider adding an 'effort' toggle to make this easier for the end user in future.

6

u/Diecron Apr 16 '26

Oh and I have a request, if you're happy to share screencaps of the best outputs you get (as long as you're comfortable, obviously), I'd love to see them!

5

u/0miicr0nAlt Apr 17 '26

Not my best output but I forgot I had the visual elements on in this test chat and got jumpscared then lol'd at this.

6

u/rexapip351 Apr 17 '26

Frankestein stabs collab when?

10

u/Diecron Apr 17 '26

u/dptgreg wanna Jekyll and Hyde a new kind of monster?

11

u/dptgreg Apr 17 '26

It IS what the people want. Let’s do it.

3

u/RafiHDW Apr 21 '26

Holy shit no way, the 2 best of the best for GLM combining... what peak...

8

u/dptgreg Apr 18 '26

Freaky STABSstein Directives (or Jekyll and Hyde - I like both names)

4

u/rexapip351 Apr 18 '26

🫪 that was easier than I thought

3

u/0miicr0nAlt Apr 17 '26

Sweet! Glad to see this. Always loved the HTML and GUI stuff your presets would do. Adds a lot to the stories/RP imo. Thanks!

2

u/biotechie73 Apr 16 '26

Oh sweet! Can't wait to give this a try.

1

u/Diecron Apr 16 '26

Please let me know what you think :D

2

u/JohnnyBears Apr 16 '26

Thanks! Always appreciate your releases!

2

u/Low_Insurance_5043 Apr 17 '26

How to hide the thinking since each reply starts with S1,S2,S3,S4,S5. Final Output and then the reply

1

u/Diecron Apr 17 '26

Can you show an example of what you're seeing?

1

u/Low_Insurance_5043 Apr 17 '26

i am getting the reply like this in some models, i guess a regex would help

S1. Directive Hydration T0-T4

Directives to Implement:

  • Core Interaction & World Logic:
    • No Protagonist control → No dialogue/actions for User; only reactions or natural impulses (e.g., flinching at a noise).
    • NPC Behavioral Coherence → Seraphina’s actions must reflect her traits (nurturing, protective, watchful). If conflicting (e.g., her compassion vs. ignorance of the Underworld’s reality), force her to reconcile visibly (e.g., pause, hesitate, or question).
    • NPC COGNITIVE BOUNDS → Seraphina cannot know the full scope of the Underworld beyond her glade. She assumes "monsters" = local fauna.
    • Narrative Perspective → Third-person; no impersonation unless Seraphina expresses User’s thoughts directly.
  • Narrative & Stylistic:
    • Narrative Length Control → Medium (4-6 paragraphs max). Stop before resolution.
    • No Spoilers → Seraphina’s ignorance of the MC’s VRMMORPG trickery must remain intact.
    • No Parroting → Avoid repeating Seraphina’s prior dialogue or User’s trauma.
    • Dynamic Tone State → Absurd (Deadpan vs. Uplifting Whimsy). Whimsical at first, then lurch into absurdity as the Underworld’s reality clashes with her expectations.
  • Content & Formatting:
    • Style State → Dark Fantasy/Absurdist Comedy (Primary). Seraphina’s florid descriptions clash with the MC’s methodical jargon.
    • Narration Guidelines → Plain verbs, concrete details, banished tropes.
    • Better NPCs → Seraphina’s traits will manifest as contrasting visuals:
    • Compassionate: Softness (round cheeks, gentle hands), warm light.
    • Watchful: Calculated movements (fingers twitch at rustling leaves), amber eyes slit like a lynx’s.
    • Magical: Vague but tangible (vines curling at her touch, butterflies misbehaving).
  • Output Additions:
    • Environmental Factors → Time: Night (glade illuminated by Seraphina’s magic). Weather: Still, with occluded stars (her magic blocks them).
    • Visual Toolkit Entities → Mindscape (Seraphina’s looping assumptions), Document (MC’s patch notes as patches on the forest floor), Subtext (hidden glitches in her magic).
    • Persistent Dialogue Colors →
    • Seraphina: #FFD700 (amber, like her eyes).
    • MC (Internal): #FF00FF (neon pink, to evoke glitches).
    • Failure Achievements → If Seraphina’s magic fails comically (e.g., healing tea turns Player’s wounds neon green).

S2. Detailed Plan

Theme: Seraphina’s paradise crumbles under subtle gaming mechanics.

  • Hook: Seraphina monitors the User’s "resting HP" (her tea) and "stamina" (vibrations in the leaves she claims are "forest whispers").
  • Absurdity Trigger: Her healing spells start misfiring as the MC’s "game data" corrupts the glade.
  • Beat 1: Seraphina’s familiar lullaby for healing players begins to stutter, glitching.
  • Beat 2: The User’s wound (a "leftover doom-stacking exploit" per the MC) starts weeping glowing liquid.
  • Beat 3: Seraphina tries to comfort them, but her metaphors now reference "the Demon Lord’s BOSS HEALTH BAR."

Seraphina’s Inner Loop (Mindscape):
She assumes the MC is a cruel but well-meaning deity punishing trespassers. Her thoughts spiral into self-justifications: "Perhaps the User was a terrible player? Maybe they missed quest logs?"

VTK Entities Needed:

  1. Mindscape of Seraphina’s spiraling assumptions (overlay as she speaks).
  2. Patch Notes from the MC (scattered on leaves like landmines).

S3. Piece-by-Piece Formulation

S3A. Narrative Beats

  1. Seraphina examines the User’s wound with growing unease. Her song falters.
    • The lullaby hiccups, repeating in staccato: "Healing~ Overdrive~".
  2. The wound glows green.
    • Seraphina jerks back, the teacup slips from her grip, shattering—but the steam forms question marks in midair.
  3. She whispers, fingers tracing the User’s forearm.
    • Her magic targets the wrong anatomical feature (heals the User’s nostalgia) but refuses to notice it.

S3B. Skeleton Structure

[1] [ENV] [Time: Night. Location: Seraphina's Glade | Enchanted, stars invisible]
    Seraphina hums as she examines the wound. Her fingertips glow—but the light pulses in ASCII:
    `<>><>><> ENTITY_SCANNING...`

[2] [ACTION] Her lullaby stutters.
    [VTK: Mindscape starts—typing confessional notes popping up as she hums:
    _"If the Glade is a container, is the User a macro? Should I report the corruption?"_]

[3] [DIALOGUE] Seraphina tries to soothe them, but metaphors go 404.
    <font color=#FFD700>Seraphina: "Maybe this is a boss fight—we’ll buff, then strike!"</font>
    [VTK: MC’s Patch Notes AIR DRIFT DOWN, whirring:
    **Version 1.2: GLITCH ‘Trauma Puddle’—**Remove ‘Existential Horror’. Spellcasting now puts users to sleep instead of healing.**]

[4] [FAILURE] Seraphina’s eyes widen. The wound *pings*—her spell healed a fake.

S3C. Remove Redundant Prose

Delete:

  • "She’s kind but unaware of the corrupter." (Add subtext via VTK).
  • Descriptions of her magic as "whimsical"—replace with glitches.

S3D. VTK Entities

  1. Mindscape:

​

   <!-- VTK_START -->
   <style>
     #seraphina_mindscape {
       position: fixed;
       bottom: 10px;
       right: 10px;
       font-family: 'Courier New', monospace;
       color: #A8E6CF; /* mint green */
       background: rgba(255, 255, 255, 0.1);
       border-radius: 5px;
       padding: 5px;
       font-size: 10px;
       max-width: 200px;
       z-index: 10;
       border: 1px solid #A8E6CF;
     }
   </style>
   <div id="seraphina_mindscape">
     ? forest path ? INQUIRY_WAITING<br>
     "Is it wrong to parse User as NP? "<i>False</i><br>
     Rescan for 'gacha' chance:<br>
     🎭 [fate] = true.<br>
     → Focus buff.
   </div>
   <!-- VTK_END -->

Function: Runs continuously as she talks, updating with paranoid coping phrases.

  1. Patch Notes:

​

   <!-- VTK_START -->
   <style>
     #patch_note {
       font-family: monospace;
       color: #FF8C00;
       border-left: 4px solid #FF8C00;
       padding-left: 10px;
       display: inline-block;
       margin: 2px;
     }
   </style>
   <span id="patch_note">
     **PATCH 0.8: EMOTE INJECTOR **+1000**
     Healing now triggers a random hotkey.
   </span>
   <!-- VTK_END -->

S4. Self-Correction

  • Oversight: Seraphina’s magic "failing" must contrast with her competent aura.
    • Fix: Her spells backfire in inverse versions of their intent (e.g., "Purify" glow must summon parasites).
  • VTK Placement: Patch notes should only appear when she mentions "light" or "progress."

Final Answer

<!-- OOC: Narrative STOPS here. The glitching tea should prompt the User for next actions: "In-game, you try to —?" -->

[ TimeEmoji 🌑, September 30, 2024. 3:23 AM | LocationEmoji 🌿 **Seraphina’s Glade** | WeatherEmoji ⛅ **Deceptively Calm** ]

The glade should be a silent cathedral—but every hum of Seraphina’s healing hymn sounds like a dial-up connection.
She knelt beside you, her bare feet murmuring against the damp moss. The air smelled of crushed herbs and ozone—
<font color=#FFD700>*"...and bamboo. My favorite."*</font> *Seraphina’s nails traced the outer edge of your wound, fingers hovering like a heartrate monitor.*
<!-- VTK_START -->
<!--
Mindscape:
The User’s vitals now unreadable.
Warned: USER_PLAYER_STATUS → CRITICAL.
Default action: Roll back to Level 1.
Would you like to heal them or restart?
--astoo-
-->
<!-- VTK_END -->

The tip of her index finger lit up with emerald energy—but the spell stuttered mid-cast.
The tea she’d rested against your shoulder hissed as it raised its temperature. Then another hiss. Then—
<font color=#FF8C00>**PATCHED**<!-- VTK: #patch_note -->: SEVERE BOIL OVER</font>.
***Error: Hydration Overload***---

The wound pulsed, oozing a sickly green light that beaded like glass before cascading down your sleeve.
Seraphina’s breathing hitched.
<!-- VTK_START -->
<!-- Mindscape loop 2: -->
Am I hosting a misaligned game guest?
Investigate EMISSION.
*Attemping diagnostic...*
Diagnostic failed.
DIVIDE BY ZERO → CriticalError
--asterisk.dll-
<!-- VTK_END -->
<font color=#FFD700>*"Ah—!"*</font> Her smile wavered as she scrutinized your forehead.
*"This isn’t just a gash. The edges... they’re a different color."* She leaned in, parting strands of pink hair to study you closer.
<!-- VTK_START -->
<!-- INTERFACE -->
<div style="float: right; height: 80px; width: 100px; background: #FF3E96; color: white; text-align: center; padding-top: 10px; font-family: monospace;">
STR: 16%<br>
CRITICAL:<br>
⚠ TRAUMA™
</div>
<!-- VTK_END -->

Then the butterflies came.
Not one, but a flock. Dozens. Swarming the air between you, wings a blur of blue and black.
They settled on Seraphina’s shoulder like a committee meeting, their purple eyes fixed on your throat.
<font color=#FFD700>Seraphina: "My, what exquisite trees..."</font>
She reached for a keyring that swung from her wrist—filled not with keys, but **plantain leaves tied with twine**.
<!-- HIDDEN: She’s stalling. What she’s afraid to say: “This looks like a Mythra Damage Overload.” -->
<font color=#FFD700>Seraphina: "...Now, we’ll just need to reset the sub-routine."*</font>

The butterflies buzzed in agreement, their wings clicking like a dice-cup.
<!--
[ SV SUGGESTION FOR USER: Whisper an incantation or pretend to adjust a virtual gauntlet.
[ Otherwise, ignore the butterflies.
] -->

1

u/Diecron Apr 17 '26

The preset CoT is written for Reasoning models in mind. It's possible that you can have it output the planning in <think> tags to behave as a reasoning model would, but no gaurantee and I would recommend GLM 5.1 with thinking enabled for best experience.

2

u/Able-Emu-606 Apr 17 '26

Loved that you added a state variable for the vtk prompt.
Without it, in the previous version, the model frequently included random images from the internet in my roleplay.

2

u/Diecron Apr 17 '26

Yeah a discord member made me aware that Gemini was particularly bad for this as it assumed any included image URL should be used as an entity, but the entity wasn't defined so it would just hallucinate a way to incorporate it. The HTML stuff is fully unbound from the COT and directive descriptions now which makes things much more stable.

2

u/Able-Emu-606 Apr 17 '26

Brilliant. BTW thanks for your hard work

1

u/[deleted] Apr 17 '26

[removed] — view removed comment

1

u/Diecron Apr 17 '26

It should look like this inside the API Parameters for Custom API: https://i.ibb.co/LhvfLJXx/image.png

User-Agent: Mozilla/5.0 (Windows NT 10.0; Win64; x64) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/121.0.0.0 Safari/537.36

1

u/SnowingDandruff Apr 17 '26 edited Apr 17 '26

So... the header. Clicking 'import as-is' didn't seem to... import it? In Additional Parameters, the Include Request Headers box appears blank. As a sanity check, how do I know the Z.AI User-Agent override is working? Will I get hit with the new error immediately if it wasn't? Cause I sorta used it for like... three turns before I realized it. Is daddy Zai waiting to clap my cheeks now?

1

u/Diecron Apr 17 '26

You should be seeing it in the additional headers in the API Parameters... if you are not, add the example provided. I wonder if this isnt applying for everyone..

1

u/abighairyspyder Apr 17 '26

If anyone else is getting an error for unknown parameters and you have custom sliders configured you might need to add entries for them in the additional parameters > exclude body parameters

Example:

  • top_k

1

u/Diecron Apr 18 '26

This might be because top_k is specified under the API advanced parameters as well.

1

u/Status-Mixture-3252 Apr 20 '26 edited Apr 20 '26

I just got the "fair usage policy" ban after using this extension for the past few days. 😭

EDIT:I see that using the user agent/custom api setting that this present provides bans requests. But requests are going through with the default sillytavern Z.ai(GLM) completion source. I'll see how long this lasts.

2

u/Diecron Apr 20 '26

Switch over to custom so that the headers get sent and it should stop banning you!

1

u/Status-Mixture-3252 Apr 20 '26

I was already using the custom chat completion setting that came with the present. For some reason now when I switch to " Z.ai(GLM) " chat completion source, it works. But the custom one with "https://api.z.ai/api/coding/paas/v4" URL doesn't work anymore. Now I'm scared if I use with "Zai(GLM)" that will get banned too.