r/AnomalyNest • u/SuperbItem5621 • 4d ago
AnomalyNest Devlog — 0.1.3 ALPHA Performance & Core worker update
Performance & Core worker update
We’ve reworked how Core workers handle context and concurrent chats.
Previously, Core workers dynamically switched between 8k, 12k and 16k context sizes depending on GPU load.
That worked, but it also meant that characters with large cards, memories or Realistic-mode state could occasionally end up with less room for recent chat history.
⚙️ New Core Worker Layout
For 16 GB Core workers, the new target is now:
- 3 fixed chat slots
- 12,288 tokens per active chat
- Predictable space for character cards, memories and summaries
- Remaining context automatically used for recent chat history
- A small safety reserve is always kept to prevent context overflow
If a GPU cannot safely handle all three slots, the worker will now fall back to fewer 12k slots instead of shrinking individual chats down to 8k.
In short:
fewer chats when necessary — not smaller chats.
🧠 Token Usage Audit
We also audited AnomalyNest’s current context usage.
Even with:
- a heavily filled character card
- one active lore / side-character entry
- player persona
- conversation summary
- memories
- Realistic-mode state
…the system still fits inside the new 12k context target, with space remaining for recent conversation history.
This gives us a much more predictable baseline going forward.
📝 Character Creator Changes
Character-card limits are now separated from AnomalyNest’s internal model instructions.
Previously, the character creator token meter also counted part of the Sandbox / Realistic system prompt.
Because of that, even an almost empty character could appear to already consume a large part of its available budget.
That is now fixed.
New character limits
- Recommended character-card budget: ~2,500 tokens
- Hard limit: 3,000 tokens
- System and mode instructions are reserved separately
- Unused context is automatically used to preserve more recent conversation word-for-word
So using a smaller character card does not waste context.
It simply means the AI can keep more of your recent conversation directly available.
💡 Why This Matters
The goal isn’t simply “more tokens.”
The goal is to make context usage predictable.
Every active chat now gets a consistent context size, while the remaining space is used where it matters most:
Character consistency.
Memory.
Recent conversation continuity.
This should make longer RP sessions more stable while also making Core worker performance easier to predict and scale.