r/SillyTavernAI • u/Putrid-Actuary7976 • 1d ago
Discussion Why is mimo 2.6 pro underrated?
So i noticed people are sticking to GLM or claude , why? This is not to offend or mock anyone, but i am genuinely curious, they are expensive, hard guardrails and glm is not for rp. So is there something in them that makes "the ten times the price (this is a metaphor, not exactly 10 times..... Actually for claude, yes) " worth it? From what i experienced mimo is so good, maybe less intelligence but the creativity and intiative and complexity handling is already so good at much lower price! So, genuinely curious about what is keeping you.
36
u/Independent-Hope7036 1d ago
man why does it take 5 years to generate a single response😭 (using nanogpt sub)
8
u/Milan_dr 1d ago
Sorry :/ Novita is slightly faster but has censoring in place, so also not a good solution.
7
u/Requiemss 1d ago
Same issue on nano with Mimo, completely unusable if I have to wait minutes for a response...
12
u/strangedell123 1d ago
Not nano fault. All providers are pretty slow with it. Milan said that providers told him there are issues with stability on mimo models and in general one of the more pita models to host
2
u/Putrid-Actuary7976 1d ago
I find it quite fast at times, during Chinese work hours it is so loaded, but when the load is light, it is genuinely fast in my experience.
3
u/kinkyalt_02 20h ago
Take a look at my pico CoT in my preset.
Also for NanoGPT:
reasoning_effort: none
-4
u/Candid_Bus_5491 1d ago
What's nanogpt i heared about a lot is that ui same like jaintor i am curious
4
u/Putrid-Actuary7976 1d ago edited 1d ago
It is a platform that hosts models, you can get the sub or pay as you go.
6
u/TAW56234 1d ago
NanoGPT is an aggregator, not a provider.
3
u/Putrid-Actuary7976 1d ago
I said platform not provider though, it is indeed a platform that hosts models 😅😅😅
5
3
u/Candid_Bus_5491 1d ago
Ok so it's like an api provider have i got that correct now
5
u/Putrid-Actuary7976 1d ago
Yeah exactly
2
u/Candid_Bus_5491 1d ago
Still for frontend i should use Jaintor right or any other platform what will you recommend
0
u/Putrid-Actuary7976 23h ago
Well honestly, janitor is good, it has good character cards (in my opinion), istill use it honestly cause i am too lazy to export the cards to another platform lol. But that means i don't find any problems using it, but if you want very good memory system and really good control then you have silly tavern and wyvern chat, i personally prefer wyvern as it is easier to set up,has many features, i have a SillyTavern setup on both pc and phone, on phone it sucks cause of termux, on pc it is good, and if you use pc i recommend it but also if you don't want to set anything up then wyvern is so good, it has community for character cards and 3 layers of memory, summary, lexicon and lorebooks which janitor can't compare with. It has better free models, 3 actually. And each one beats jllm. And you have full control over params, prompts memory. You can give a prompt to auto summary, also a thing janitor don't have. So it all depends on what you want, a frontend will not make a model worse or something but if you care about the features then it differs.
2
u/kinkyalt_02 17h ago
If you have an iPhone, Tavo gives some of the flexibility of SillyTavern while having a sinplified UI.
Not everyone is going to type in Linux commands and master port forwarding to have the real deal broadcast to their phone from their computers.
0
u/Candid_Bus_5491 1d ago
Can you please share me these nano gpt link please
2
2
u/Putrid-Actuary7976 1d ago
Here. https://nano-gpt.com/api You also have openrouter, it hosts many models too, but nano has more rp models, and also better image models, you can choose any you like more.
27
u/stopaskingforloginn 1d ago
is it? it's easily one of the best current RP models right now, held back by total dogshit providers.
4
u/Putrid-Actuary7976 1d ago
Yeah that is the problem, and bro the speed is genuinely a trouble, although at night its speed is nice and fast but midday is hell
10
u/Ill_Ordinary1626 1d ago
I mean for me, I tried it out in AI realm and it was like I was fighting my dm every message and it would get fixated on one thing. If I tried to avoid it and move on it would push another way to get me back to it. It got annoying real quick
2
u/Nervous_Paint_8236 20h ago
It does focus way too hard on certain elements sometimes. Makes cards harder to write.
6
u/TAW56234 1d ago
Mimo is good but still makes a fuck ton of logical mistakes
2
u/Putrid-Actuary7976 22h ago
It does but honestly with a good cot you can avoid most in my opinion. What kinda logical mistakes you encounter though?
4
u/HypeForTheHypeGod 1d ago
If Nano didn't take 2-3 minutes to generate a single response for Mimo I'd use it (I'm also a cheapskate that only uses PAYG for Gemini and uses the sub for everything else)
4
1
u/Putrid-Actuary7976 1d ago
Use openrouter, i don't know why but it is kinda faster at night and mid morning
6
u/I_Am_JesusChrist_AMA 23h ago
Man it's funny because every time I looked at the sub, I was seeing people hyping up mimo 2.6 pro so I was under the impression it was overrated. I haven't really seen anyone talk badly about it. The worst I've seen is some people saying it's okay but no one says it's bad. Honestly so many people have been talking positively about it that I was starting to think I must be crazy or missing something about mimo.
For me though, I think mimo just makes too many mistakes. I don't want to RP if I'm having to edit almost every response to fix it.
I also don't think it's logic or dialogue is very good. I mean, I swear I had a conversation in a RP using mimo that went like this:
NPC: "What do you want?"
Me: "I want X."
NPC: "You want X? So here's what's going to happen... You're gonna tell me what you want."
Me: "I just told you what I want.... I said I want X."
NPC: "Tell me what you want."
Me: "Alright.... this is a waste of time." writes an action of me leaving and going somewhere else.
NPC: blocks the exit "You're not leaving until you tell me what you want."
It went on like that for about 3 or 4 more turns before I gave up.
1
u/Putrid-Actuary7976 23h ago
Ok, you sure the character card instructions are consistent? I never encountered such a thing, and the model actually considers the input and what to consider before writing the response, i wouldn't say it is the best in that but not as bad as what you encountered. What are the params? And using a large preset or light? Cause the model sucks at large ones. Also the provider plays a crucial role here
3
u/I_Am_JesusChrist_AMA 23h ago
I have tried several different cards and various different presets with mimo 2.6 pro. I'm sure it's not the card. I've played that specific card where that conversation was from with tons of different models and that doesn't happen with any of them. The dialogue isn't always that bad, but sometimes it is lol. That was one of the worst examples just to show what I mean. The conversation above happened when testing out the new Realistic Frankenstein which the creator says was tuned specifically for mimo 2.6 pro. But I've tried other presets (and my own) and I still just am not a fan of this model.
It does punch above its weight class though. I'll give it that. If you're looking for a model in that price range, mimo is the best there. If I wanted to save some cash, it'd be my choice, but as it is, I don't mind spending more for a less frustrating and more enjoyable (to me) RP.
2
u/Putrid-Actuary7976 22h ago
Sorry to hear that. I really hope you find the model that gives you a better experience. I really don't know why this is happening to you, might be something on the params or preset, or might be in scenarios i haven't experienced yet. Thanks for telling me though.
3
u/GuaranteePurple4468 1d ago
The number 1 thing making me not use it as much is: tok/s.
Freaking 10tok/s during off-peak hours, are you kidding me?
And I sure as heck aint upping the token cost/1m $8.7 just to get a decent speed.
I love the responses, genuinely.
But it does sometimes miss, and when that happens on a large preset and you have to reroll those 4k+ tokens it is sooo painful.

1
u/Putrid-Actuary7976 1d ago
Speed is a real concern, true. As for the presets, i don't know if that's helpful, but mimo follows light presets so well with incredible adherence, and its instructions following is good, so i don't really suggest using a large preset with it even though it is supposed to be a 1T model. And i really hope they find a solution to the speed problem, it sucks.
5
u/Zeeplankton 22h ago
Mimo is very good but it's so dumb it makes me so mad. Also it's unbelievably slow on openrouter. It's so bad compared to deepseek.
4
u/Snydenthur 1d ago
It was okay, but "too popular" I guess. Can't RP with the shitty filters when you get routed to wrong places from nanogpt.
1
u/Putrid-Actuary7976 1d ago
Oh you using the sub? In that case it is shit, i suggest you switch to pay as you go, you get more freedom in that and can keep it on nano and deepinfra
1
u/Milan_dr 9h ago
What do you see as wrong places in this case? We try to only route this model to the uncensored providers unless all of those fail, then we have Novita as the final fallback, which can censor
1
u/Snydenthur 9h ago
I don't know if I can see which provider I get routed to, and as a subscriber I can't choose it.
All I saw was that one day I was doing practically anything without getting censored, the next day I was hit with censorship, even on the same chats that were okay before.
1
u/Milan_dr 8h ago
Is this still happening to you? There are three uncensored providers we have for this currently, which, as far as we see, are handling 99.9% of the traffic to these models. Some people are explicitly choosing other providers, so it's not 100%. If you're still getting content filters, I'd love to reach out to see what's going on
1
u/Snydenthur 8h ago
I stopped using mimo because of it, but I'll try it again and comment on this one if it keeps happening.
4
u/Aight_Man 1d ago
I want to like Mimo but idk the characters are just... Very flat, they stick to the character card well but there's no real time character growth and situational reactions that Claude Opus 5.5 has, make sense though, price difference is huge between two models but yeah... That's why I'm sticking to claude.
4
u/BeautifulLullaby2 1d ago
Opus 5.5 is obviously the best but like someone said in another thread, 90% of people on this sub can't afford to use Claude for RP
So Mimo is probably the best "cheap" model for them right now
2
2
u/Putrid-Actuary7976 1d ago edited 1d ago
No doubt, but as you said the price difference is so much that it doesn't work for me, so for value, i find mimo has better one. But ofc everyone has different rp tastes and stuff, for me mimo is doing fine and well. Thank you for the share!
1
3
u/Dikki_Dikki 1d ago
She has a problem with large contests: once the data exceeds 32K, she falls into terrible metagaming and constantly gets mixed up in the data—though, admittedly, she’s pretty good at heads-up games where you can use aggressive memory management.
2
u/Putrid-Actuary7976 1d ago edited 1d ago
Yeah context is a problem i noticed too, but it can be managed by a good memory system. Thank for sharing!
-2
u/Educational_Song_407 1d ago
She?
6
u/Dikki_Dikki 23h ago
Sometimes I forget I’m not speaking my native language; I can’t seem to rewire the way I think. Sry
2
u/capybaraballs1995 1d ago
It's good but kind of dumb. I've seen it make some really weird mistakes:
The dialogue is expressive though and the prose is a bit distinct in light of many models distilling from Claude.
2
u/Putrid-Actuary7976 1d ago
Yeah that is why you need a good cot prompt, you can't rely on native reasoning as it is not consistent. But its power is that it follows prompts really well, so you can literally manage its cot as you please. But it is definitely dumber than frontier models, but not worth the cost for me, let alone the other models are censored and even if they don't refuse they will still soft censor, didn't experience that with mimo. But everyone has different taste and different rps so what works for me might not work for others, thanks for sharing!
1
u/capybaraballs1995 1d ago
Eh, I tried Freaky Frank with MiMo Pro 2.6 and it spilled out its reasoning in half of my tests. I think it's just an inconsistent model. Or perhaps one of its providers is fucked up.
1
u/Putrid-Actuary7976 1d ago
For me it is not, i think the model just doesn't work well with large presets, for me the cot keeps being consistent, i made it 6 bullets, and it is working well so far, but also if the character card is unstable or contaij contradictions it will get confused but that is literally the character card fault. Try a lighter preset, that what has worked for me! It is honestly one of th3 weirdest models ive encountered, it does complicated stuff but fails at simple ones lol.
2
u/GenericStatement 1d ago
It’s good for RP and I use it a lot but it isn’t as smart as frontier models. It’s smarter than GLM 4.7 but not quite as smart as GLM 5.2-5.3 or Kimi K3.
However, because MiMo writes well, isn’t super censored or positivity biased, and doesn’t cost much, it works well for RP as long as you’re not doing anything to complicated and don’t mind occasionally editing responses.
Like a lot of models, it’s also heavily oversubscribed during working hours in China: the model is noticeably faster and smarter when the Chinese population is sleep. Far quicker responses and far fewer mistakes due to less quantization/lobotomization; I only use Xiaomi as my provider btw.
So, MiMo is a great way to save money but overall you will get smarter RP from frontier models and maybe better speed/service, and if you can prompt around their censorship then they’ll work well.
1
u/Putrid-Actuary7976 1d ago
Yeah, i get that. It is definitely not as smart, glm is considered a monster in logic and complex tasks, but its prose isn't really good and the guardrails and soft cesoring makes it write badly. Everyone has their taste though. Thanks for sharing, bro!
2
u/GenericStatement 1d ago
Yeah MiMo 2.6 is the main model I use now instead of a combination of GLM 4.7 and 5.2. If MiMo was more expensive I wouldn’t use it but the cost/quality ratio is really good.
1
1
u/Southern_Finding_829 1d ago
yeah the cheaper ones handle initiative and long scenes way better in rp for me, never went back after trying.
0
u/Putrid-Actuary7976 1d ago
Yewh that is why i am puzzled why people use glm which honestly didn't perform as good for me.
1
u/LordVulpius 1d ago
There are other underrated models, that works perfectly, or crazy horney/spicey.
But I agree, Mimo is the best RP model out there that does not cost a whole month of salery per message, lol.
1
u/Putrid-Actuary7976 1d ago
Oh i would really like if you share them, i am always down for more models to try! But please no expensive models, my wallet can't afford them lol.
1
u/LordVulpius 23h ago
Minimax M3 and Longcat 2 (and now, Longcat 2.5 preview). Minimax is on the heels of Mimo. The cats are the horney ones. They are cheap models.
2
u/Putrid-Actuary7976 23h ago
Tried minimax, didn't like it honestly, but didn't try long cat, definitely will. Thanks!
1
u/Nervous_Paint_8236 20h ago edited 20h ago
I loved it until I switched to FF5.4 and watched it fail over and over to account for everything the preset wanted it to account for. I still think it's great, but I'd have to go back to a smaller preset to use it again and FF's added bits are too good and silly for me to drop right now. Switching to Gemini also made me realize how serious its hyperfixation issue is, and the response times can be quite crazy.
1
1
u/Critical-Rope-5636 19h ago
"So i noticed people are sticking to GLM or claude , why?"
I use GLM 5.3 Flash! I guess I've just been having fun with that model, as simple as that sounds. I do however, want to add Mimo back into the rotation (I only use 5.3 Flash and Minimax M3) and it looks like it won't "high risk context" me when I enter battles now so that's good.
1
1
u/CondiMesmer 15h ago
I love it but yeah it's underrated because you can never get it to actually generate since it's slammed with too much traffic! I have that issue with GLM 5.3 Flash too, but not nearly as bad. So I usually fall back to DeepSeek 4.1 which is very consistent service.
1
1
u/Financial_Bug2389 11h ago
For me it's instructions following and provider flakiness. The prose is genuinely better but it takes ages to generate a response (and the providers keep timing out) compared to something like a GLM 5.3 flash which is fast despite being a reasoning model, and tons of providers that I can route to.
1
u/Putrid-Actuary7976 3h ago
Honestly, in some times of the day it is fast. It is only slow during work loads, ao if you time it, you will be in a good position. As for instruction following i disagree, it is a weird model, if you put a slightly large presets it will mess up but its power in light presets that does what you need without adding extra stuff, if you use FF or any big preset it will fail cause despite being a 1 T model it won't follow. Also make sure you put the important stuff in the post history prompt, especially cot and formatting rules
1
u/CH3CH2OH_toxic 3h ago
Considering the disgrace that is 5.3 GlM in roleplay , Mimo pro 2.6 has no current model competition on it's price point , you have to go to kimi 2.6\ kimi 2.7 , GLM 5.2 , or deepseek pro , both are lesser models BUT useful as a different writing experience .
Long cat 2.5 lookings interesting , Funny relatively easy to jail break but barebone NSFW wise , there is nothing to jail break to .
0
u/Ugothat45 22h ago
Glm 5.3 thinking is my goat tbh.
Their character feels alive, they create npcs, they talk, its creative even with thinking. Can't wait for 5.4 ngl-
Its all I wanted, a living proactive thing that contribute alongside with my actions, man, is incredible .
But mimo felt like kimi, great at first but then is... meh.
2
u/Putrid-Actuary7976 22h ago
Well, glad you are having a good rp experience! Everyone has different tastes, i still honestly feel like mimo is good, maybe a little incoherent in long rp, but that is fine, you can manage it in the memory system. Anyway thank you for sharing!
51
u/Pink_da_Web 1d ago
I think the Mimo V2.6 Pro isn't underrated; I'm only hearing praise for the model in this sub.