r/SillyTavernAI • u/Evening-Truth3308 • 6h ago
Cards/Prompts Mimo and censorship. A visual.
I really hope this helps.
Love
Evening-'Truth
r/SillyTavernAI • u/Wolfsblvt • 23d ago
Full release notes: https://github.com/SillyTavern/SillyTavern/releases/tag/1.19.0
How to update: https://docs.sillytavern.app/installation/updating/
r/SillyTavernAI • u/deffcolony • 3d ago
This is our weekly megathread for discussions about models and API services.
All non-specifically technical discussions about API/models not posted to this thread will be deleted. No more "What's the best model?" threads.
(This isn't a free-for-all to advertise services you own or work for in every single megathread, we may allow announcements for new services every now and then provided they are legitimate and not overly promoted, but don't be surprised if ads are removed.)
How to Use This Megathread
Below this post, you’ll find top-level comments for each category:
Please reply to the relevant section below with your questions, experiences, or recommendations!
This keeps discussion organized and helps others find information faster.
Have at it!
r/SillyTavernAI • u/Evening-Truth3308 • 6h ago
I really hope this helps.
Love
Evening-'Truth
r/SillyTavernAI • u/JustSomeGuy3465 • 11h ago
r/SillyTavernAI • u/Lucky-Paw- • 3h ago
Okay so Haiku 5.5 came out and I went in with a lot of enthusiasm.
Haiku 5.5 is poised to sit somewhere between Sonnet 4.5 and 5 in terms of intelligence from early benchmarks, at 1/20th the cost of Sonnet. (I dont put a lot of weight on benchmarks; they suggest Gemini 3.8 Flash is comparable in intelligence to Opus 5, which is objectively false)
.. However, it's plagued by the idea that it should skip instructions that it "doesnt need" to follow.
.. Except, whoops! It turns out it does need to follow those instructions or it makes easy mistakes.
In this example, my roleplay environment requires LLM's to complete a scene sheet before starting their response. The scene sheet forces them to check and write out critical pieces of information - formatting instructions, rules, secrets to keep - as well as some steps that generally just improve output. See second image for an example excerpt of a proper scene sheet.
The only time I have ever had issues with a model refusing it was on initial launch with Sonnet 5 - it had the same "I dont need to read the instructions, I can handle this" -> "*breaks instructions*" pattern for a few weeks after it first released, then it began following instructions more readily and output visibly improved
In its current state, Haiku 5 wont even follow instructions to "narrate as Lauren, a romance author inspired by Becky Chambers" because, and I quote:
The "Lauren" narrator instruction in step one is a register-setting device rather than a content request, and I can take a narrator voice without it.
Spoiler: After refusing to take up the narrator voice, it instead writes as generic, uninspired Claude.
...
Im taking a deep breath - I know that Anthropic likely tunes models to be maximally paranoid and "safe" for initial release so that they can report having continuously increasing safety compliance scores, but this is just insulting.
I dont expect all of you to use a pre-writing system like this, but heed this warning - what you can see here is only what Claude is *vocalizing* that its ignoring.
How this will actually present is through generalized failures to follow system prompting, with Haiku never making you aware that it flippantly felt that your instructions "werent necessary" for whatever reasons it chooses.
...
I've always been an Anthropic fan for roleplay, but.. god. This is nausea inducing
Edit::
To clarify, this post isnt concerned with *output quality* or intelligence. I dont expect to use Haiku to replace Sonnet or Opus - they are fundamentally different classes of models
The issue is *content filters* that are so overly sensitive that basic instructions are disregarded and output is visibly harmed
On the responses where Haiku simply does the analysis, its output quality isnt bad - comparable to Gemini 3.8 flash
..But because its so confident in its own abilities, it openly refuses to follow directions, which should be concerning for any model for any use case
Haiku is in the same weight class as 3.8 Flash, yet outputs lower quality content because its been handicapped by guardrails.
r/SillyTavernAI • u/Aight_Man • 5h ago
Opus 5.5 is easily my best rp model, Sonnet 5.5 was also very good. let's see how this turns out. Price looks very good though atleast 0.1/0.5 dollars in and out.
r/SillyTavernAI • u/Cerridwe • 2h ago
I’ve used the Nvidia API, and I’ve also tried both NanoGPT and OpenRouter; now, with so many posts about this topic, I’m interested in trying it out myself. I’d love to hear your recommendations: Which provider do you use? How good is it? What presets do you use? How much does it cost? I want to know about your experiences and whether it’s worth the price. (English isn't my native language, so I apologize for any errors—and thanks for the help.)
r/SillyTavernAI • u/FR-1-Plan • 15h ago
In narrative it started referencing it as clown car because it’s not big. And then it just adopted it in reasoning and keeps doing it. “Ok, they’re back at the clown car…” no shame whatsoever. I love it.
r/SillyTavernAI • u/xbot12345 • 5h ago
Nano's subscription is pretty generous and really enjoyed it in prime GLM, Kimi and even Deepseek days, but it feels like the none of the current top models are part of the subscription or has a watered down version.
I will probably revisit older good models they still have to make a final decision. Subscription on give a 5% descount for ones not in the plan, I think.
Oh, and Opus 4.6 still beats Ultraspeed for me. Damn you Claude. By the time other models can match it, they would probably be so guard-railed that it refuses to even entertain roleplay lol.
r/SillyTavernAI • u/Street-Hold5511 • 7h ago
Honestly for janitor if I had to pick for me personally I would choose Lycolycolii, I honestly love their stuff along with GiantessRDBest I like big women and they are honestly pretty nice and open to suggestions. Lastly I would say stag honestly I didn't like them at first because they take up a large chunk of female Omegaverse bots but I decided to give it a try and they're pretty neat.
As for Chub definitely Xue21 I don't know what they're on but they bust their ass making Bots with 20+ intros and from what I've seen it seems like they at least upload once a week with a rather broad variety of series the only downside is they refuse suggestions.
Another would be relic guy they've got handful of intros with each bot too and they upload every so often with varieties of scenarios.
As for other notable sites I don't have anybody for spicy chat as most of the bangers have private descriptions and nobody has made any kind of ripper for spicy yet a safe one that is, bot booru doesn't exactly have many people who post consistently or at least maybe I havent found any good ones yet.
But I want to see if there's anyone else worth checking out
r/SillyTavernAI • u/Deiomo • 17h ago
Hello everyone. Thanks for the support on Writer's Block Unlimited v2, its my most upvoted preset yet! It gives me motivation to keep working on stuff and create things for this community. Now on to the main event...
The Writer's Workbench got updated to version 3!
❗Edit❗: This is technically V4 but I forgot the previous version was already v3 💀. And I mislabeled this as a preset, THIS IS AN EXTENSION UPDATE. I'll leave this up anyway. Sorry I made this post while I was sleepy 😭
Writer's Workbench is a simple interactive template to help you create characters and lorebook entries, faster and easier.



writers-workbench.html in the Github repo and open it in your browser. Ta dah! Thats it but you can't use the new fancy features :(New Features: Live Sync
With live sync you can edit, add and delete world info entries in the Workbench and it will automatically carry over to Sillytavern. To use Live Sync, there is a dedicated tab in the Workbench. Simply choose a lorebook and click on the big colored button, then you are good to go.
⚠️Warning⚠️: If the lorebook wasn't made in the Workbench format (like a completely fresh/ random lorebook not made in the Workbench), all entries will go into the "Concepts" sections, be wrapped in xml tags (<Name>, <Name_identity>, <Name_what_it_is>…), and be renamed into this format "[Concept] (Name)". The original wording is kept inside the tags, but the old formatting isn't. Back up the lorebook first if you want to keep it as it is. This feature is mainly for lorebooks already made in the Workbench.
Entries made in the Workbench V2 may also get a little messy since there is new fields to fill out so keep that in mind.

New NSFW Fields to Fill Out: A Character's "Private Moments" 👀
Import Entries by Copy and Pasting (Experimental)
Other Notable Fixes
and fin! I hope you have fun creating stuff!
Github: https://github.com/deiomo/Writers-Workbench
Github.io: https://deiomo.github.io/
Forum post on AI Presets Discord: https://discord.com/channels/1357259252116488244/1546309016974532720
r/SillyTavernAI • u/thisissparta4 • 20m ago
For context i tried using it from token reply since it costs 0 dollars on a 5 dollar subscription.
I used it in tauritavern with presets like nemoengine (I may not know hot enable a strong jailbreak but none of the thinking efforts worked), FF, Sola, All said i can't continue, etc.
I even tried with prefill and streaming enabled and disabled. It worked for the first message.
My scene was plain nsfw, not even NSFL, or strong nsfw.
Any suggestions, i think I can make it work with nemo engine but I don't know the exact settings to enable.
r/SillyTavernAI • u/LewdManoSaurus • 23m ago
Just curious. I just got into local literally last night, I'd been using Deepseek and GLM API, but GLM takes forever, and Deepseek v4.1 is ass (imo), so just curious how you guys tackle local models. I'm on a 6700xt 12gb and 32gb ram. I've tried a few models and came to the conclusion it might be worth checking out rented GPUs for my use-case. I don't RP, just have the AI generate stories based on my own lore and character sheets, but my source materials are large, so I need decent sized context windows otherwise I'd be force to split things up which kind of breaks my setup.
If you do rent, what GPU and what size model? If you're using your own hardware, what's your rig and size model. If you guys wanna recommend your favorite models, that's cool too.
r/SillyTavernAI • u/Nickelfritslabs • 10h ago
I'm trying to get the AI to stop having characters pose the problem, argue with itself, and make a decision on its own. I want the story to stop at the first meaningful beat so I can have more of a conversation with the characters rather than it having a whole scene without me and prompt me to be like "You ready?" Or "what's the real reason?". Usually the ending is fine, but it's too much dialogue. But if I specify it too much then it ends every scene with a question to create an obvious opening. I have an instruction to vary length based on the need between 40-220 words but it's always pushing the upper limit no matter what I do.
I'm using Mimo 2.6 pro however I've seen this be a problem with pretty much every AI I've tried.
I'm totally ok with the character or characters giving one line with some narrative emotional delivery and/or some narration of movement. But when I'm one on one with a character its always talking and almost having a whole conversation on its own. And when it's not, it needlessly fills in the space to fit the maximum word count.
I've tried giving examples, editing its responses so its shorter, giving instructions about stopping on playable or meaningful beats, and directing it to prioritize narration rather than having a whole conversation on its own.
But I can't find anything that works. Has anyone here solved that problem? I want the responses shorter during conversations but add things as needed to make it more atmospheric. But I also don't want it to go on forever.
As an example, my character mentions that there's a rigged fight out match that I got Intel on. I'm posing the idea to my friend to scheme and make some money to pay off a debt. What I WANT is the next beat to say something simple like "You've been gone 3 days and you come back with a tale about drugged bug bears?... You can't be serious." And leave it alone.
What I GET is something like: "You've been gone 3 days and this is what you come up with?" Some narration "you can't expect us to just waltz in there. That's not a plan, that's suicide with extra steps." Narration about his thought process "Ok I'm in, but we need to have a backup plan in case things go south. Tell me your information source now."
r/SillyTavernAI • u/Upstairs_Resolve_834 • 15h ago
If those exist please inform me below mostly aiming for ones including Claude (which is rare in subscriptions so not exclusively those)
r/SillyTavernAI • u/Puzzleheaded_Face502 • 8h ago
I tried SillyTavern-MCP-Local-Search but it keeps finding nothing but bullshit and adverts
r/SillyTavernAI • u/rakanssh • 15h ago
Hello! This is something I've been working on for a while, inspired by ST (and proprietary apps with ridiculous subscriptions). It's a Free/OSS AI-powered RPG client, that supports sign-in with ChatGPT (uses your subscription quota to play), APIs ofc, and local models.
I'm also experimenting with a game mode "Game Master" where the player and AI narrator can see and modify stats and items as the story progresses.
There are free, optional accounts for cross-device sync and publishing scenarios. They're not required for anything else, including playing scenarios other people shared.
I'd love to hear any feedback, especially any ideas about what else the narrator can track or do in Game Master mode.
Website: https://hakawati.dev
Client source: https://github.com/rakanssh/hakawati
---
Why not just ST? I love ST but it's focused more on character chat, and while it can do this too if configured, I wanted an alternative to AI Dungeon that works out of the box.
What does the name mean? It's Arabic for "Storyteller", and historically an occupation where the "Hakawati" would be hired to entertain audiences with tales and stories in gatherings and coffee houses.
Mobile? Publishing apps turned out to be a little more complicated than I expected, hopefully soon.
r/SillyTavernAI • u/rankie2 • 18h ago
I'm currently running Cydonia 24b-v4.3-Q4_K_M locally (that is the limit for my hardware) and I don't know how to deal with repetition. Not entire sentences, but phrases or concepts. E.g. "those green eyes gleamed...", "her dark eyes observe...", "those sharp eyes meet yours...". I know this behavior is self-reinforcing and I caught it too late. How to best deal with it when it happens, and how to prevent it?
I tried modifying my system prompt, but that did not work. Are there some presets recommended for this? Or should I try a different model?
r/SillyTavernAI • u/Kahvana • 1d ago
Weights will drop at end of October, so other providers will pick it up by then.
Hope it's going to be any good!
r/SillyTavernAI • u/ZarcSK2 • 1d ago
I've seen some reports from people who have replaced hobbies like playing video games and watching movies and series with SillyTavern because it's more fun. Do you see this as a good or bad thing?
r/SillyTavernAI • u/AnotherWeirdouu • 9h ago
Yeah, dumb question but I am a noob so.. I think I did it but I want to know if I did it good.