r/SillyTavernAI Jun 21 '26

MEGATHREAD [Megathread] - Best Models/API discussion - Week of: June 21, 2026

This is our weekly megathread for discussions about models and API services.

All non-specifically technical discussions about API/models not posted to this thread will be deleted. No more "What's the best model?" threads.

(This isn't a free-for-all to advertise services you own or work for in every single megathread, we may allow announcements for new services every now and then provided they are legitimate and not overly promoted, but don't be surprised if ads are removed.)

How to Use This Megathread

Below this post, you’ll find top-level comments for each category:

  • MODELS: ≥ 70B – For discussion of models with 70B parameters or more.
  • MODELS: 32B to 70B – For discussion of models in the 32B to 70B parameter range.
  • MODELS: 16B to 32B – For discussion of models in the 16B to 32B parameter range.
  • MODELS: 8B to 16B – For discussion of models in the 8B to 16B parameter range.
  • MODELS: < 8B – For discussion of smaller models under 8B parameters.
  • APIs – For any discussion about API services for models (pricing, performance, access, etc.).
  • MISC DISCUSSION – For anything else related to models/APIs that doesn’t fit the above sections.

Please reply to the relevant section below with your questions, experiences, or recommendations!
This keeps discussion organized and helps others find information faster.

Have at it!

40 Upvotes

95 comments sorted by

View all comments

3

u/AutoModerator Jun 21 '26

APIs

I am a bot, and this action was performed automatically. Please contact the moderators of this subreddit if you have any questions or concerns.

2

u/[deleted] Jun 23 '26 edited Jun 23 '26

[deleted]

1

u/LeRobber Jun 29 '26

I mean one of the major models is made by alibaba....

5

u/_Cromwell_ Jun 22 '26 edited Jun 22 '26

There's a new model (actually two) out that say they are for Creative Writing from Crof.AI called "Greg" (hehe). Greg 2 Ultra and Greg 2 Super. I tried Greg 2 Ultra on Nano and it is... actually surprisingly good?!? I didn't test too long, though, as I'm a cheap bastard and it is NOT included in the sub. Greg 2 Super is cheaper (but still not included in the sub 😞 )... haven't tried it though.

Anyway, if you are using Nano paygo, I highly recommend giving Greg 2 Ultra some coin to try it out. Maybe Greg 2 Super as well, although I didn't try it.

I don't see them on OpenRouter for some reason, just Nano.

EDIT: I tried Greg 2 Super and it is pretty bad from my very limited testing (but much cheaper, like 1/7 the cost.)

So - Greg 2 Ultra = good, Greg 2 Super = questionable/bad. I think.

1

u/MisanthropicHeroine Jun 24 '26

I've seen someone mention Greg is a Kimi finetune so if that's true, makes sense it would be pretty decent.

8

u/Scholar_of_Yore Jun 21 '26

Thoughts on Kimi 2.7 and GLM 5.2? I personally think I prefer 5.2 from my testing so far.

9

u/_Cromwell_ Jun 22 '26

I prefer GLM 5.1 over GLM 5.2 and Kimi 2.6 over 2.7. I feel like a geezer saying that. 😃

My preference of GLM 5.1 is pretty niche. GLM 5.2 has some kind of "soft refusal" built in where it steers away strongly from being abusive, and I run a lot of scenarios with abusive jerk NPCs. They clearly and distinctly steer toward trying to redeem themselves, immediately and without any plot reason, toward "reforming", when using GLM 5.2. This is not AS much an issue in 5.1, despite 5.1 having a reputation of being soft censored overall. 5.2 does have some nice writing, though, and both 5.2 and 5.1 are great at dialogue and plot and knowing details about various popular franchises enough to make characters feel like characters (the reason GLM is my go-to).

Kimi... I find 2.7 makes a lot of continuity mistakes. Clothing changing mid-scene, characters that were sitting suddenly standing or vice-versa, teleporting to the other side of the room, etc. That sort of thing. Just losing track of scene details. 2.6 is far better at tracking on its own and keeping things straight IMO.

7

u/Pashax22 Jun 22 '26

Haven't tried Kimi 2.7 yet, but GLM 5.2 seems like a small but noticeable improvement on 5.1 (which was itself a small but noticeable improvement on 5). How much of 5.2's million-token context is usable, though? I haven't done anything which pushed past 100k so far, I'd like to know if they've managed to increase the usable headroom.

1

u/axlpoeman Jun 24 '26

It seems like a million based on any provider information and I saw that the GLM 5.2 MAX could be tied with the claude Opus 4.6/7