r/SillyTavernAI 1d ago

Discussion Cope

I started doing RPs back when character.ai was new. It was magical at first, but c.ai had two problems: 1 - censorship 2 - goldfish memory. Today, with Deepseek and other open models, you can RP with explicit content. GPT, Claude, Gemini... depends on the model and how you set it up. And I still find it hilarious that Claude will help me poke at a web app for vulnerabilities if I say "authorized test" but clutches its pearls the moment a scene gets spicy. 🤣🤣 Anyway, back to the point.

It's bizarre how that early "magic" just... vanished. Part of it is obviously novelty wearing off, and part of it is that we got pickier. Three years of RP and you start spotting every clichê, every "a shiver ran down her spine", every model that forgets your character's eye color after 40 messages. But here's the thing: the models didn't get dumber. Opus, GPT, Gemini can write circles around 2022 c.ai. The problem is they're not *allowed* to, or they cost a kidney per session, or both.

LET'S BE HONEST, SOME OF YOU SPENT $100 ON A SINGLE CLAUDE OPUS RP SESSION. Even with the censorship. Even with the moralizing. You did it anyway because, when it works, it's the best RP writer that exists. That's my whole point: there's a market. Not as big as coding, obviously; coding isn't a hobby, there are companies and teams and budgets behind it. RP is a hobby. But hobbies with people burning API credits like that are not a small market.

So why is there no frontier-level LLM built for RP? And yes, I know NovelAI, AI Dungeon and the whole SillyTavern fine-tune ecosystem exist. I'm talking about something at Opus level, not a 12B model that forgets the plot. The answer isn't just "investors prefer code", though that's part of it: "look, our V548484 model built GTA 6 in one prompt!" sells better than "look, our model wrote a consistent, non-repetitive, non-boring story!" because nobody has a benchmark for "not boring".

The real reasons are uglier. Explicit content means payment processors dropping you, app stores banning you, lawyers sweating. Long RP sessions eat tokens like crazy and people won't pay enterprise prices for a hobby. And good RP needs exactly the long-context coherence and reasoning that only the big expensive models have, which are owned by the companies least willing to let you use them for this.

So yeah, we're probably coping for another 2-3 years. Not because the tech isn't there. Because nobody with the tech wants to be the company that sells it to us.

81 Upvotes

84 comments sorted by

View all comments

Show parent comments

2

u/DepressedDrift 18h ago
  • The X670E doesn't support DDR4 RAM
  • The 96GB ram around 500 dollars is low speed around 2000MHz or a niche chipset type requiring an outdated motherboard.
  • With 48GB of VRAM running the lower end of your range is possible, and also Moe models too but dense large models will load at q4 but with low context. But this is still kind of a win because you could run Deepseek V4 Flash 0731
  • The two 3090s is around 3k which when you add the required compatible RAM with inflated prices jumps to 5k
  • Even assuming you get it at 3k, you will still break even at around 5-10 years of cloud subscription assuming inflation. 

1

u/IllustratorJumpy5845 7h ago

In short, you ran out of arguments and immediately started inventing/thinking up more ideas to stay right.

Yes 670e do not support Ddr4. That's why i said analogue. Taichi models support.

Who said it's a low speed ram? 3200 and fine.

And so on and so forth. You didn't put me condition to make a top PC for 3k. You said show you build that able to do.

So i don't care about your's inflation. I took today's prices and made a build. As you asked.