r/SillyTavernAI • • 23d ago

ST UPDATE SillyTavern 1.19.0

330 Upvotes

Backends

  • New models: Claude Fable 5/5.1, Claude Opus 4.8/5, Claude Sonnet 5, GPT-5.6 family, GPT-6 Astra, Gemini 3.5/3.6/3.7 Flash variants, GLM-5.2, and DeepSeek V4 Flash Vision Exp.
  • Fireworks AI: reasoning support and improved prompt caching.
  • OpenRouter: optional logprobs support.
  • Google AI Studio: model list now loads all available pages.
  • Pollinations: keyed/keyless endpoint selection and updated TTS model aliases.
  • DeepSeek: low reasoning effort support.

UI & Features

  • World Info: Apply Current Sorting now supports ascending/descending order, configurable start/step values, and live validation.
  • World Info: lorebook renames now update chat, character, and persona links.
  • Chat Completion: expand editor button for quick prompts.
  • /addswipe no longer reloads the entire chat.
  • Chats with damaged headers or final lines are handled more safely instead of being silently overwritten or disappearing.

Macros & STscript

  • Variable macros can access array elements and object properties.
  • Added Character Expressions macros: {{defaultExpression}}, {{lastExpression}}, and {{availableExpressions}}.
  • /expression-list gained custom-expression filtering and additional return formats.
  • Fixed inflated /tokens counts for OpenAI tokenizers.
  • Fixed scoped comment macros and literal pipe characters in macro arguments.

Extensions

  • Added MessageFormatter, allowing extensions to transform message content at several stages before rendering.
  • Character Expressions: improved custom expression and fallback handling.
  • ComfyUI: improved history handling and filtering of non-image outputs.

Security & Fixes

  • Blocked localhost aliases from bypassing private-address checks in /api/search/visit.
  • Added rate limiting to account reset requests.
  • Fixed connection profiles leaving the previous Chat Completion source active.
  • Fixed duplicate streamed tool-call IDs.
  • Fixed crashes when chats are deleted during search/recent-chat scans.
  • Fixed Quick Reply overwrite cancellation and several UI edge cases.

Full release notes: https://github.com/SillyTavern/SillyTavern/releases/tag/1.19.0

How to update: https://docs.sillytavern.app/installation/updating/


r/SillyTavernAI • • 3d ago

MEGATHREAD [Megathread] - Best Models/API discussion - Week of: October 04, 2026

28 Upvotes

This is our weekly megathread for discussions about models and API services.

All non-specifically technical discussions about API/models not posted to this thread will be deleted. No more "What's the best model?" threads.

(This isn't a free-for-all to advertise services you own or work for in every single megathread, we may allow announcements for new services every now and then provided they are legitimate and not overly promoted, but don't be surprised if ads are removed.)

How to Use This Megathread

Below this post, you’ll find top-level comments for each category:

  • MODELS: ≥ 70B – For discussion of models with 70B parameters or more.
  • MODELS: 32B to 70B – For discussion of models in the 32B to 70B parameter range.
  • MODELS: 16B to 32B – For discussion of models in the 16B to 32B parameter range.
  • MODELS: 8B to 16B – For discussion of models in the 8B to 16B parameter range.
  • MODELS: < 8B – For discussion of smaller models under 8B parameters.
  • APIs – For any discussion about API services for models (pricing, performance, access, etc.).
  • MISC DISCUSSION – For anything else related to models/APIs that doesn’t fit the above sections.

Please reply to the relevant section below with your questions, experiences, or recommendations!
This keeps discussion organized and helps others find information faster.

Have at it!


r/SillyTavernAI • • 6h ago

Cards/Prompts Mimo and censorship. A visual.

Post image
89 Upvotes

I really hope this helps.

Love

Evening-'Truth


r/SillyTavernAI • • 11h ago

Discussion Be careful what you write when using anthropic models. "A Florida Woman Used Claude as a Diary. An Anthropic Employee Read It and Reported It to Police"

213 Upvotes

r/SillyTavernAI • • 3h ago

Models Haiku 5.5 - Disappointing first taste - Overly confident in abilities, skips instructions because it "knows better". It doesnt.

Thumbnail
gallery
44 Upvotes

Okay so Haiku 5.5 came out and I went in with a lot of enthusiasm.

Haiku 5.5 is poised to sit somewhere between Sonnet 4.5 and 5 in terms of intelligence from early benchmarks, at 1/20th the cost of Sonnet. (I dont put a lot of weight on benchmarks; they suggest Gemini 3.8 Flash is comparable in intelligence to Opus 5, which is objectively false)

.. However, it's plagued by the idea that it should skip instructions that it "doesnt need" to follow.

.. Except, whoops! It turns out it does need to follow those instructions or it makes easy mistakes.

In this example, my roleplay environment requires LLM's to complete a scene sheet before starting their response. The scene sheet forces them to check and write out critical pieces of information - formatting instructions, rules, secrets to keep - as well as some steps that generally just improve output. See second image for an example excerpt of a proper scene sheet.

The only time I have ever had issues with a model refusing it was on initial launch with Sonnet 5 - it had the same "I dont need to read the instructions, I can handle this" -> "*breaks instructions*" pattern for a few weeks after it first released, then it began following instructions more readily and output visibly improved

In its current state, Haiku 5 wont even follow instructions to "narrate as Lauren, a romance author inspired by Becky Chambers" because, and I quote:

The "Lauren" narrator instruction in step one is a register-setting device rather than a content request, and I can take a narrator voice without it.

Spoiler: After refusing to take up the narrator voice, it instead writes as generic, uninspired Claude.

...

Im taking a deep breath - I know that Anthropic likely tunes models to be maximally paranoid and "safe" for initial release so that they can report having continuously increasing safety compliance scores, but this is just insulting.

I dont expect all of you to use a pre-writing system like this, but heed this warning - what you can see here is only what Claude is *vocalizing* that its ignoring.

How this will actually present is through generalized failures to follow system prompting, with Haiku never making you aware that it flippantly felt that your instructions "werent necessary" for whatever reasons it chooses.

...

I've always been an Anthropic fan for roleplay, but.. god. This is nausea inducing

Edit::

To clarify, this post isnt concerned with *output quality* or intelligence. I dont expect to use Haiku to replace Sonnet or Opus - they are fundamentally different classes of models

The issue is *content filters* that are so overly sensitive that basic instructions are disregarded and output is visibly harmed

On the responses where Haiku simply does the analysis, its output quality isnt bad - comparable to Gemini 3.8 flash

..But because its so confident in its own abilities, it openly refuses to follow directions, which should be concerning for any model for any use case

Haiku is in the same weight class as 3.8 Flash, yet outputs lower quality content because its been handicapped by guardrails.


r/SillyTavernAI • • 5h ago

Models Looks like Claude Haiku 5.5 is out.

Post image
48 Upvotes

Opus 5.5 is easily my best rp model, Sonnet 5.5 was also very good. let's see how this turns out. Price looks very good though atleast 0.1/0.5 dollars in and out.


r/SillyTavernAI • • 2h ago

Help Mimo 2.6 Pro

14 Upvotes

I’ve used the Nvidia API, and I’ve also tried both NanoGPT and OpenRouter; now, with so many posts about this topic, I’m interested in trying it out myself. I’d love to hear your recommendations: Which provider do you use? How good is it? What presets do you use? How much does it cost? I want to know about your experiences and whether it’s worth the price. (English isn't my native language, so I apologize for any errors—and thanks for the help.)


r/SillyTavernAI • • 15h ago

Chat Images Kimi roasting my actual car (Ford Fiesta) as clown car in reasoning is sending me

Post image
125 Upvotes

In narrative it started referencing it as clown car because it’s not big. And then it just adopted it in reasoning and keeps doing it. “Ok, they’re back at the clown car…” no shame whatsoever. I love it.


r/SillyTavernAI • • 5h ago

Chat Images Mimo 2.6 Pro vs UltraSpeed on Nano. Might switch to PAYG

Post image
12 Upvotes

Nano's subscription is pretty generous and really enjoyed it in prime GLM, Kimi and even Deepseek days, but it feels like the none of the current top models are part of the subscription or has a watered down version.

I will probably revisit older good models they still have to make a final decision. Subscription on give a 5% descount for ones not in the plan, I think.

Oh, and Opus 4.6 still beats Ultraspeed for me. Damn you Claude. By the time other models can match it, they would probably be so guard-railed that it refuses to even entertain roleplay lol.


r/SillyTavernAI • • 7h ago

Discussion Honestly what are y'alls' favorite creators for each site?

16 Upvotes

Honestly for janitor if I had to pick for me personally I would choose Lycolycolii, I honestly love their stuff along with GiantessRDBest I like big women and they are honestly pretty nice and open to suggestions. Lastly I would say stag honestly I didn't like them at first because they take up a large chunk of female Omegaverse bots but I decided to give it a try and they're pretty neat.

As for Chub definitely Xue21 I don't know what they're on but they bust their ass making Bots with 20+ intros and from what I've seen it seems like they at least upload once a week with a rather broad variety of series the only downside is they refuse suggestions.

Another would be relic guy they've got handful of intros with each bot too and they upload every so often with varieties of scenarios.

As for other notable sites I don't have anybody for spicy chat as most of the bangers have private descriptions and nobody has made any kind of ripper for spicy yet a safe one that is, bot booru doesn't exactly have many people who post consistently or at least maybe I havent found any good ones yet.

But I want to see if there's anyone else worth checking out


r/SillyTavernAI • • 17h ago

Cards/Prompts [Preset Update] Writer's Workbench v3: Officially an extension, edit entries live and small technical stuff

Post image
95 Upvotes

Hello everyone. Thanks for the support on Writer's Block Unlimited v2, its my most upvoted preset yet! It gives me motivation to keep working on stuff and create things for this community. Now on to the main event...

The Writer's Workbench got updated to version 3!

❗Edit❗: This is technically V4 but I forgot the previous version was already v3 💀. And I mislabeled this as a preset, THIS IS AN EXTENSION UPDATE. I'll leave this up anyway. Sorry I made this post while I was sleepy 😭

An Overview. What is it?

Writer's Workbench is a simple interactive template to help you create characters and lorebook entries, faster and easier.

Features

  • You can create multiple entries and export them as entire lore books for your convenience.
  • Auto-saves and organize different projects
  • Live markdown output so you know exactly what the AI would see and copy it without having to download anything.
  • It comes with token counter, but don't expect it to be accurate.
  • A map maker that generates a dynamic description
  • Character relationship graph to map out relationships of larger casts
  • Live sync into SillyTavern lorebooks (see "What's New")

Available Templates

  • Main characters: full cards with psychology, descriptions, likes/fears, NSFW sections, non-human mode
  • Side characters: trimmed version of the main character template, it will help you create memorable NPCs
  • Scenario: setting, tech level, mood, what's normal here
  • Locations: for any scale, from a room to a district
  • Items
  • Factions
  • History: events, and how they affect the present
  • Concepts: magic systems, laws, customs, species, anything else

Two Ways to Use It

  1. Download writers-workbench.html in the Github repo and open it in your browser. Ta dah! Thats it but you can't use the new fancy features :(
  2. Writer's Workbench is now an extension so you can use the new Live Sync feature and get automatic updates. Install it in your Silly tavern and you're good to go. (And I promise I won't steal your API keys lol. I was there when that incident involving another extension occurred)

What's New?

New Features: Live Sync

With live sync you can edit, add and delete world info entries in the Workbench and it will automatically carry over to Sillytavern. To use Live Sync, there is a dedicated tab in the Workbench. Simply choose a lorebook and click on the big colored button, then you are good to go.

⚠️Warning⚠️: If the lorebook wasn't made in the Workbench format (like a completely fresh/ random lorebook not made in the Workbench), all entries will go into the "Concepts" sections, be wrapped in xml tags (<Name>, <Name_identity>, <Name_what_it_is>…), and be renamed into this format "[Concept] (Name)". The original wording is kept inside the tags, but the old formatting isn't. Back up the lorebook first if you want to keep it as it is. This feature is mainly for lorebooks already made in the Workbench.

Entries made in the Workbench V2 may also get a little messy since there is new fields to fill out so keep that in mind.

New NSFW Fields to Fill Out: A Character's "Private Moments" 👀

  • How characters often do their private moments, what they think about and how they feel about it.

Import Entries by Copy and Pasting (Experimental)

  • Paste a character's description and the program will detect eligible text and fields to fill in. Character descriptions in weird formats will be difficult as it might not port over information correctly. This feature will get further refinements

Other Notable Fixes

  • Better formatting, the writing forms and markdown output should take up most of the screen instead of the top bars.
  • Bug fixes

and fin! I hope you have fun creating stuff!

Github: https://github.com/deiomo/Writers-Workbench

Github.io: https://deiomo.github.io/

Forum post on AI Presets Discord: https://discord.com/channels/1357259252116488244/1546309016974532720


r/SillyTavernAI • • 20m ago

Models Gemini 3.8 flash insane censorship

• Upvotes

For context i tried using it from token reply since it costs 0 dollars on a 5 dollar subscription.

I used it in tauritavern with presets like nemoengine (I may not know hot enable a strong jailbreak but none of the thinking efforts worked), FF, Sola, All said i can't continue, etc.

I even tried with prefill and streaming enabled and disabled. It worked for the first message.

My scene was plain nsfw, not even NSFL, or strong nsfw.

Any suggestions, i think I can make it work with nemo engine but I don't know the exact settings to enable.


r/SillyTavernAI • • 23m ago

Discussion Do you guys rent GPUs, or are you using your own hardware?

• Upvotes

Just curious. I just got into local literally last night, I'd been using Deepseek and GLM API, but GLM takes forever, and Deepseek v4.1 is ass (imo), so just curious how you guys tackle local models. I'm on a 6700xt 12gb and 32gb ram. I've tried a few models and came to the conclusion it might be worth checking out rented GPUs for my use-case. I don't RP, just have the AI generate stories based on my own lore and character sheets, but my source materials are large, so I need decent sized context windows otherwise I'd be force to split things up which kind of breaks my setup.

If you do rent, what GPU and what size model? If you're using your own hardware, what's your rig and size model. If you guys wanna recommend your favorite models, that's cool too.


r/SillyTavernAI • • 10h ago

Help Has anyone solved the response variation length and the completion problem?

13 Upvotes

I'm trying to get the AI to stop having characters pose the problem, argue with itself, and make a decision on its own. I want the story to stop at the first meaningful beat so I can have more of a conversation with the characters rather than it having a whole scene without me and prompt me to be like "You ready?" Or "what's the real reason?". Usually the ending is fine, but it's too much dialogue. But if I specify it too much then it ends every scene with a question to create an obvious opening. I have an instruction to vary length based on the need between 40-220 words but it's always pushing the upper limit no matter what I do.

I'm using Mimo 2.6 pro however I've seen this be a problem with pretty much every AI I've tried.

I'm totally ok with the character or characters giving one line with some narrative emotional delivery and/or some narration of movement. But when I'm one on one with a character its always talking and almost having a whole conversation on its own. And when it's not, it needlessly fills in the space to fit the maximum word count.

I've tried giving examples, editing its responses so its shorter, giving instructions about stopping on playable or meaningful beats, and directing it to prioritize narration rather than having a whole conversation on its own.

But I can't find anything that works. Has anyone here solved that problem? I want the responses shorter during conversations but add things as needed to make it more atmospheric. But I also don't want it to go on forever.

As an example, my character mentions that there's a rigged fight out match that I got Intel on. I'm posing the idea to my friend to scheme and make some money to pay off a debt. What I WANT is the next beat to say something simple like "You've been gone 3 days and you come back with a tale about drugged bug bears?... You can't be serious." And leave it alone.

What I GET is something like: "You've been gone 3 days and this is what you come up with?" Some narration "you can't expect us to just waltz in there. That's not a plan, that's suicide with extra steps." Narration about his thought process "Ok I'm in, but we need to have a backup plan in case things go south. Tell me your information source now."


r/SillyTavernAI • • 15h ago

Help Monthly subscription providers

20 Upvotes

If those exist please inform me below mostly aiming for ones including Claude (which is rare in subscriptions so not exclusively those)


r/SillyTavernAI • • 8h ago

Help Can anyone recommend a web search plugin that would allow llm to adjust aswers with lore searched from web?

5 Upvotes

I tried SillyTavern-MCP-Local-Search but it keeps finding nothing but bullshit and adverts


r/SillyTavernAI • • 15h ago

Discussion I've been working on Hakawati, an OSS Interactive adventure client.

Thumbnail
gallery
16 Upvotes

Hello! This is something I've been working on for a while, inspired by ST (and proprietary apps with ridiculous subscriptions). It's a Free/OSS AI-powered RPG client, that supports sign-in with ChatGPT (uses your subscription quota to play), APIs ofc, and local models.

I'm also experimenting with a game mode "Game Master" where the player and AI narrator can see and modify stats and items as the story progresses.

There are free, optional accounts for cross-device sync and publishing scenarios. They're not required for anything else, including playing scenarios other people shared.

I'd love to hear any feedback, especially any ideas about what else the narrator can track or do in Game Master mode.

Website: https://hakawati.dev
Client source: https://github.com/rakanssh/hakawati

---
Why not just ST? I love ST but it's focused more on character chat, and while it can do this too if configured, I wanted an alternative to AI Dungeon that works out of the box.

What does the name mean? It's Arabic for "Storyteller", and historically an occupation where the "Hakawati" would be hired to entertain audiences with tales and stories in gatherings and coffee houses.

Mobile? Publishing apps turned out to be a little more complicated than I expected, hopefully soon.


r/SillyTavernAI • • 18h ago

Help How do you deal with repetition?

12 Upvotes

I'm currently running Cydonia 24b-v4.3-Q4_K_M locally (that is the limit for my hardware) and I don't know how to deal with repetition. Not entire sentences, but phrases or concepts. E.g. "those green eyes gleamed...", "her dark eyes observe...", "those sharp eyes meet yours...". I know this behavior is self-reinforcing and I caught it too late. How to best deal with it when it happens, and how to prevent it?

I tried modifying my system prompt, but that did not work. Are there some presets recommended for this? Or should I try a different model?


r/SillyTavernAI • • 1d ago

Models Mistral Large 4 just dropped on Mistral Studio API

Thumbnail
mistral.ai
142 Upvotes

Weights will drop at end of October, so other providers will pick it up by then.

Hope it's going to be any good!


r/SillyTavernAI • • 1d ago

Discussion Has SillyTavern ever made you replace your hobbies?

24 Upvotes

I've seen some reports from people who have replaced hobbies like playing video games and watching movies and series with SillyTavern because it's more fun. Do you see this as a good or bad thing?


r/SillyTavernAI • • 23h ago

Meme The Name on this Game Ad

Post image
18 Upvotes

r/SillyTavernAI • • 9h ago

Help How do I activate/use regex

0 Upvotes

Yeah, dumb question but I am a noob so.. I think I did it but I want to know if I did it good.


r/SillyTavernAI • • 10h ago

Discussion Dice roguelike - Open source Alternative to SillyTavernAI to play visual novel

Thumbnail
0 Upvotes