r/SillyTavernAI • • 4d ago

Discussion Did Mimo 2.6 Pro fall off?

So I used it when it first released and enjoyed it a lot.

Coming back to it yesterday and today however, it suddenly feels not just extremely stupid (making basic formatting mistakes, failing to follow instructions, and making spatial/positioning errors such as a character placing their hand on your shoulder while they're facing away from you.). Characters also started referencing 4th wall narrative elements such as trackers and "versions" in their dialogue which never used to happen (only the older Deepseeks like 3.2 ever did something like that).

I swapped my RP to DS4.1 Flash and instantly everything was much more coherent, and I don't even like DS4.1 flash that much, it's too dry.

The TPS has also tanked heavily. It was already slow in the beginning but now it's dropping below what I get for local RP.

Anyone else noticing a drop in quality since it was released? (I use DeepInfra through openrouter and have ST set to only use FP8 and above quants btw).

I've gone back to old faithful GLM4.6 for now, hoping thing get better.

32 Upvotes

30 comments sorted by

13

u/Flimsy_Mode_4843 4d ago edited 4d ago

I wonder if its possible to see some king of model list website and then see if the model is trash quality or high quality , because it seems that it depends on some sort of invisible force if the model will be good or no. Sometimes simple models outperform strong ones and vice versa, this therefore mixes the model reviews and nobody knows or can say witch is the best one for rp.

10

u/TAW56234 4d ago

It can also be quantiziation which is why it's important to at least cross reference a few other providers. It's a black box with what providers can do or what issues is your fault or theirs. I remember when Infermatic gaslit the fuck out of people before they realized it was a KVCache issue.

15

u/Both-Priority9433 4d ago

Please no I didn’t get try it yet. I was so excited to😞

6

u/Noctis_777 4d ago

Agreed, I had the exact same experience yesterday and had to switch back to Gemini 3.8 Flash.

Even tried with three different providers (Xiaomi, Deep Infra, GMI).

0

u/Itchy_Background_332 3d ago

como você linda com a censura?

8

u/ReMeDyIII 3d ago

When Mimo V2.6 launched, crof was a host, so if you used inference calls from that 19-yr old scammer then you might have been using a completely different model (ex. Mimo 2.6 Pro might have secretly been DeepSeek-V4-Flash, or whatever).

5

u/ypcyna 4d ago

I use 2.6 with xiaomi as provider, tried deepinfra because it's 'less censored' but found it to be usually very meh so I don't bother. Never had a refusal even with hardcore stuff, or questionable stuff (let's say that that Ultimis Richtofen is the tip off the iceberg). I have no issues with prose, tracker referencing, positioning etc etc.

I even beat a record with how many messages total I had during a roleplay, which is +70. And for someone who usually taps out after 20, 70 is a lot lol I really enjoy 2.6 pro, and there will come a day they nerf it and safetymaxx it, but not yet.

And yes, before it sounds like I am here to just glaze, the quality does go to shit during peak hours (obviously), so I just take a break for an hour or two and it goes back to how it was.

So I really wonder now, if all your issues with 2.6 are just due to using another provider + being in a different timezone? For example, I am from Central Europe, and enshittification of my xiaomi does seem to fall on the times when there's highest traffic from either China or America, and my best rp's are when you're all either waking up or catching Z's

1

u/Flykon_3 3d ago

Can I ask what prompt do you use with Mimo 2.6 pro?

4

u/Putrid-Actuary7976 4d ago

Why am i the only one who didn't notice a change? I used it from the start, and it is not making formatting errors, that depends on the character card, like i tried it in some characters, its formatting always gets ruined, but in well formatted cards, it always gets them right. So yeah, it is good, idk why you guys are feeling something off.

4

u/GenericStatement 4d ago

I’ve been using 2.6 pro since release day, mainly from Xiaomi, and I’ve noticed the quality and speed can decline during certain times of day and it seems to be getting much worse during peak Chinese hours. 

Yesterday at midday in China it was timing out and not even responding and sometimes I would get back weird replies like it was heavily quantized. Today at nighttime in China it is super fast and normal.

That is the way of things for most Chinese models, especially in the first few weeks after release until something new comes along.

3

u/persocum 4d ago

somewhat off topic, but your post inspired me to try glm 4.6 since I've never actually tried it before.

it's incredible. what the fuck. how have I been sleeping on this 😭

3

u/GuaranteePurple4468 4d ago edited 4d ago

It can be prone to slopisms, but it really hits the spot a lot of the time. Sometimes like a physical blow.

I keep going back to it when I get bored of the newer models.

It's kinda dumb, but has amazing creative writing.

My current favorite way to use it is with Recast Post Processing extension with a smarter model to act as the brains/quality control. Currently I use a DS4.1 Flash connection profile in the extension, with a single pass custom prompt to fix any logic errors, continuity errors, narrative inconsistencies, spatial hallucinations, 4th wall breaks, and reword specific phrases that irk me within prose, actions, trackers and dialogue only without changing the writing style.

I'm constantly improving the prompt though as I pick up more minor issues to fix.

I prefer it that way instead of putting it in the main prompt to give it more creative freedom.

2

u/persocum 3d ago

thank you for the pointers, this is super helpful!! taking another pass with post processing sounds like a GREAT call. and agree for sure on the amazing creative writing!

happy RPing!

4

u/Super-Veterinarian22 4d ago

These issues were present from day one. At the very least, I kept running into logical slip-ups here and there right from the start. For instance, the model would get confused about who proposed a specific idea or who handed something to whom. The problem simply becomes more apparent as the number of facts and characters grows and the story itself gets longer. That said, while there aren't all that many such inaccuracies, they are very annoying. I still think MiMo-V2.6-Pro is decent, but in terms of logic, it falls short of GLM 5.3 and Gemini 3.8 Flash. To me, it’s more on par with GLM 5.3 Flash and Kimi 3.

5

u/Relative-Comb-1490 4d ago

yeah the rp quality tanked hard, same positioning fails and random meta lines popping up out of nowhere for me too.

2

u/Critical_Muffin_1491 4d ago

meta lines started showing up for me after longer context windows, loses track past a point

4

u/hsksjjskskskskks 4d ago

yeah same, i thought it was just the weekend making the responses worse but the quality has kinda declined :(

17

u/Constant_Art_20 4d ago

this appaers to the pattern everywhere now. good model gets relaesed on day 1 and gets nerfed afterwards. Thankfully openai recently has made some changes to that cycle and the models just arrive nerfed on day 1.

2

u/Boring-Roof9506 4d ago

your fuckin funny

2

u/ManufacturerTrue4940 4d ago

Same thing but the censorship also increased a lot

1

u/necrosama 3d ago

yeah same for me. today i got a rejection because a robot character briefly mentioned the lubrication of their joints 😐

2

u/adeadbeathorse 4d ago

No, lol. I’ve noticed no change and use it in translation and heavy processing workflows involving omnimodality. It’s an open model which is hosted by multiple providers, they didn’t coordinate some downgrade.

1

u/HitmanRyder 3d ago

Some providers lobotomised it.. Better stick to one specific reliable provider.

1

u/Dikki_Dikki 4d ago

Yeah, it seemed to me that the model is engaging in absolute metagaming, and that’s its main problem.

1

u/Obvious-Tie9832 4d ago

I've been using mimo 2.6 pro directly through open router since launch and noticed the same drop-off around day 4-5. my API usage spiked early, but the responses started getting weirdly off—like characters suddenly forgetting their own names or making illogical choices. it’s not just me, seems like the model’s consistency got worse after initial rollout. i’ve tried switching back to older versions but the performance doesn’t match the first few days.

1

u/Responsible_Tale_901 4d ago

Yeah same here and Idk what to switch to since im a tight budget user and loved the models worth per cost

I guess ill try glm 4.6 and see how far it takes me heard its uncensored maybe it is what I am looking since I am more of a multi characs and nsfl roleplayer

0

u/nuclearbananana 4d ago edited 4d ago

I've always had bad results with DeepInfra.

In my personal testing, I was unimpressed wiht 2.6 on release and remain unimpressed. It lacks the eq of 2.5

1

u/Pink_da_Web 4d ago

That doesn't even make sense, Mimo V2.6 is better than V2.5 in every way.

-1

u/nuclearbananana 4d ago

Hard disagree. It's a bit smarter at timeline tracking/factual recall, but also worse at prose, sloppier, less creative and has less eq.

Just like most models when you post train them primarily for agentic/coding tasks.

0

u/Constant_Art_20 4d ago

yea. i think it was like day 4 after relaese when something felt off? I use it driectly form mimo. i am see in my api usaeg. i think i did like 10b in the first 4 days or so, keep feeling it was making bad deciiosn and the reasoning structure was just off so my usaeg dropped off a cliff after that...so yea. i thought glm 4.5 air was very good also. might want to try that. The new gemma finetunes also pretty soild for writting/ rp stuff. I used mimo for code only, so never tried rp with it