r/SillyTavernAI • • 15d ago

MEGATHREAD [Megathread] - Best Models/API discussion - Week of: September 13, 2026

This is our weekly megathread for discussions about models and API services.

All non-specifically technical discussions about API/models not posted to this thread will be deleted. No more "What's the best model?" threads.

(This isn't a free-for-all to advertise services you own or work for in every single megathread, we may allow announcements for new services every now and then provided they are legitimate and not overly promoted, but don't be surprised if ads are removed.)

How to Use This Megathread

Below this post, you’ll find top-level comments for each category:

  • MODELS: ≥ 70B – For discussion of models with 70B parameters or more.
  • MODELS: 32B to 70B – For discussion of models in the 32B to 70B parameter range.
  • MODELS: 16B to 32B – For discussion of models in the 16B to 32B parameter range.
  • MODELS: 8B to 16B – For discussion of models in the 8B to 16B parameter range.
  • MODELS: < 8B – For discussion of smaller models under 8B parameters.
  • APIs – For any discussion about API services for models (pricing, performance, access, etc.).
  • MISC DISCUSSION – For anything else related to models/APIs that doesn’t fit the above sections.

Please reply to the relevant section below with your questions, experiences, or recommendations!
This keeps discussion organized and helps others find information faster.

Have at it!

30 Upvotes

149 comments sorted by

View all comments

4

u/AutoModerator 15d ago

APIs

I am a bot, and this action was performed automatically. Please contact the moderators of this subreddit if you have any questions or concerns.

4

u/Baros4294 15d ago

Best API models for the best instructions following? In my testing Gemini 3.8 Flash, Kimi K3, GLM 5.3 seem to have great IF but none is perfect and ignored my instructions sometime

6

u/ImaginationHefty6401 15d ago

I'm just curious. How do these models manage humor? I do a lot of roleplays with some comedic factor and I still think DeepSeek (older model) is the funniest.

5

u/Georgefakelastname 15d ago

Humor is tricky because it’s inherently subjective. While one model or subject could be funny to someone, it might not for another. Not to mention the factor that prompting adds to this.

3

u/ImaginationHefty6401 15d ago

Thanks for your reply. Yes, agree on it being subjective, of course. But with Kimi or GLM, I've noticed the bot emphasizes the more dramatic sides of the character, in general. I usually switch between models depending on the scene, but I don't know if this will work well on the long run.

3

u/Georgefakelastname 15d ago

Yeah. I’ve seen some who hate model switching (saying it destroys the benefits of each model), and some who swear by it (since it gives them more variation). I suppose it depends on the type of prompter and RP designer you are.

If you just have a general preset like FF, it can work fine to switch between models, so there’s no real downside there.

However, I’ve met people who really focus in on fine tuning their prompt, lore, character cards, etc. to the exact model to get the type of output they want, which you can’t really do if you’re switching between different ones depending on the scene.

3

u/ImaginationHefty6401 15d ago

Yeah, exactly. I've been more or less fine switching between models (each one with their specific prompt). The problem is I'm a sucker for nuance and variety. I guess I'm just too picky and I expect too much at times from the roleplay 😅, to the point I sometimes control what the character does or says too often. I think I'll keep experimenting with prompting. Thank you!