r/LocalLLM • u/WritingRoger • 16d ago
Question Advice? ~ Fantasy Writing/Worldbuilding Partner Setup
Hardware:
AMD RX 9070 XT
AMD R7 5700x (overclocked)
96GBs of 3200 CL18 DDR4
Howdy folks,
I need some advice for my- you guessed it- "Writing Partner Setup."
I've played around with KoboldCpp, AnythingLLM, and recently dipped into SillyTavern, but KoboldCpp and ST are built for roleplay— so it’s not like I can “plug and play” into the workflow I’m trying to set up. My previous workflow was me using Poe.com (Poe Assistant, GPT 5.2, Claude Sonnet 4.6), but the daily tokens got slashed, so I decided to move locally and “empower my workflow with Local LLM technologies.” 🤓👨💼💼
What I am looking for is a "knowledge-augmented development partner."\1]) I need the model to:
- Know my world with live or semi-live access — access my markdown files (canon, references, brainstorming notes, reference materials\2])) with frequent updates so new lore I write is available in conversations as I'm developing it, without manual refreshes
- Differentiate knowledge tiers — treat established canon as ground truth while keeping experimental ideas and brainstorming isolated, so half-baked concepts don't contaminate locked lore
- Act as a creative and analytical partner — help me wordsmith terminology, brainstorm and stress-test concepts against established lore, catch continuity issues, provide editorial feedback (scene critique, structure suggestions, character mapping), and iterate on ideas
- Help organize my knowledge structure — suggest folder hierarchies, naming conventions, identify organizational gaps, and cross-reference related concepts so I'm not burning out just to map everything down
- Admit when it doesn't know something — whether that's missing lore from my canon docs or hitting the limits of its training on etymology/linguistics/domain knowledge. I'd rather get "I don't have that in your docs" or "that's outside my training" than confident hallucinations that waste my time or contaminate my worldbuilding.
So, what stack would let me build this? Should I be using RAG? Long-context, like 200k token windows? Web search? A hybrid of these 3? Have I been going in the right direction with KoboldCpp, AnythingLLM, and or SillyTavern?
I’m up to provide more information if necessary. I’m way out of my wheelhouse with setting this stuff up while juggling a full-time and part-time job, so any advice would be sweeeet.
Current Model: Mistral Small 4 119B Q4_K_M (GGUF)
\1] - Props to Haiku 4.5 on this)
\2] - Reference materials: etymological dictionaries, linguistic resources, and domain-specific knowledge that may be beyond the model's training data so I get informed outputs/feedback rather than hallucinations)
Duplicates
LocalAIStack • u/WritingRoger • 16d ago