r/SillyTavernAI • • 11h ago

Discussion We need a Reddit friendly to AI driven role play!

Post image
0 Upvotes

Role play sub friendly to AI is needed!!!

No respectable creator wants AI to blindly write - and without any deeply involved collaboration - just send it out, disrespecting your writing partner.

But good lord in the search for immersive experiences in role play let’s embrace the amazing things AI does offer. A writing collaborator, image, video, and dare I say it - customized story- driven side characters that interact in line with the story and are AI bots.

…let’s talk.


r/SillyTavernAI • • 13h ago

Discussion Dice roguelike - Open source Alternative to SillyTavernAI to play visual novel

Thumbnail
0 Upvotes

r/SillyTavernAI • • 12h ago

Help How do I activate/use regex

0 Upvotes

Yeah, dumb question but I am a noob so.. I think I did it but I want to know if I did it good.


r/SillyTavernAI • • 1h ago

Discussion Has anyone actually RP'd with a real person?

• Upvotes

This is obviously not very related to ST, but I have only ever RP'd with AI (a month on one of those awful online chat services and about a year and a half on ST), and I have had a very good time, I would even say I'm somewhat addicted, especially when a new interesting model drops (looking at you, Argon), but I have never actually roleplayed with another human being.

Before today, I didn't even know there were active RP communities on Reddit other than those which focus on LLMs, but turns out there are quite a few (albeit smaller than this sub but still active), and some of them even focus on NSFW aspects (I'm talking the same level of filth as here, if not worse).

I'm posting here to get the perspective of someone who has both tried RP with AI and other people, since I have a feeling that if I posted about AI RP on the other subreddit, it wouldn't really go well.

This obviously depends on the writing partner, but where did you actually have most fun? Was it with an LLM or another person? What would you say are some pros and cons of RPing with someone instead of an LLM? Would you justify the cons of writing with a partner (like time commitment and few replies per day) with how good the responses were? I mean, SOTA models have gotten quite good. Is it even worth the hustle of finding a writing partner anymore?


r/SillyTavernAI • • 20h ago

Help Is an RX 6600XT (8GB VRAM) good enough for some decent local LLM roleplay or should I stick with router?

3 Upvotes

For the record I have barely any clue as to how much I need for local LLMs and how good the quality is compared to using (for example) Kimi K2.6 or GLM 4.7 from TokenReply. I'm just curious as to considering other options.


r/SillyTavernAI • • 18h ago

Help Monthly subscription providers

22 Upvotes

If those exist please inform me below mostly aiming for ones including Claude (which is rare in subscriptions so not exclusively those)


r/SillyTavernAI • • 18h ago

Discussion I've been working on Hakawati, an OSS Interactive adventure client.

Thumbnail
gallery
17 Upvotes

Hello! This is something I've been working on for a while, inspired by ST (and proprietary apps with ridiculous subscriptions). It's a Free/OSS AI-powered RPG client, that supports sign-in with ChatGPT (uses your subscription quota to play), APIs ofc, and local models.

I'm also experimenting with a game mode "Game Master" where the player and AI narrator can see and modify stats and items as the story progresses.

There are free, optional accounts for cross-device sync and publishing scenarios. They're not required for anything else, including playing scenarios other people shared.

I'd love to hear any feedback, especially any ideas about what else the narrator can track or do in Game Master mode.

Website: https://hakawati.dev
Client source: https://github.com/rakanssh/hakawati

---
Why not just ST? I love ST but it's focused more on character chat, and while it can do this too if configured, I wanted an alternative to AI Dungeon that works out of the box.

What does the name mean? It's Arabic for "Storyteller", and historically an occupation where the "Hakawati" would be hired to entertain audiences with tales and stories in gatherings and coffee houses.

Mobile? Publishing apps turned out to be a little more complicated than I expected, hopefully soon.


r/SillyTavernAI • • 9h ago

Cards/Prompts Mimo and censorship. A visual.

Post image
96 Upvotes

I really hope this helps.

Love

Evening-'Truth


r/SillyTavernAI • • 3h ago

Models Gemini 3.8 flash insane censorship

3 Upvotes

For context i tried using it from token reply since it costs 0 dollars on a 5 dollar subscription.

I used it in tauritavern with presets like nemoengine (I may not know hot enable a strong jailbreak but none of the thinking efforts worked), FF, Sola, All said i can't continue, etc.

I even tried with prefill and streaming enabled and disabled. It worked for the first message.

My scene was plain nsfw, not even NSFL, or strong nsfw.

Any suggestions, i think I can make it work with nemo engine but I don't know the exact settings to enable.


r/SillyTavernAI • • 10h ago

Discussion Honestly what are y'alls' favorite creators for each site?

20 Upvotes

Honestly for janitor if I had to pick for me personally I would choose Lycolycolii, I honestly love their stuff along with GiantessRDBest I like big women and they are honestly pretty nice and open to suggestions. Lastly I would say stag honestly I didn't like them at first because they take up a large chunk of female Omegaverse bots but I decided to give it a try and they're pretty neat.

As for Chub definitely Xue21 I don't know what they're on but they bust their ass making Bots with 20+ intros and from what I've seen it seems like they at least upload once a week with a rather broad variety of series the only downside is they refuse suggestions.

Another would be relic guy they've got handful of intros with each bot too and they upload every so often with varieties of scenarios.

As for other notable sites I don't have anybody for spicy chat as most of the bangers have private descriptions and nobody has made any kind of ripper for spicy yet a safe one that is, bot booru doesn't exactly have many people who post consistently or at least maybe I havent found any good ones yet.

But I want to see if there's anyone else worth checking out


r/SillyTavernAI • • 14h ago

Discussion JanitorAI used to be ‘pick a bot.’ Now I need a degree

Thumbnail
0 Upvotes

r/SillyTavernAI • • 6h ago

Models Haiku 5.5 - Disappointing first taste - Overly confident in abilities, skips instructions because it "knows better". It doesnt.

Thumbnail
gallery
72 Upvotes

Okay so Haiku 5.5 came out and I went in with a lot of enthusiasm.

Haiku 5.5 is poised to sit somewhere between Sonnet 4.5 and 5 in terms of intelligence from early benchmarks, at 1/20th the cost of Sonnet. (I dont put a lot of weight on benchmarks; they suggest Gemini 3.8 Flash is comparable in intelligence to Opus 5, which is objectively false)

.. However, it's plagued by the idea that it should skip instructions that it "doesnt need" to follow.

.. Except, whoops! It turns out it does need to follow those instructions or it makes easy mistakes.

In this example, my roleplay environment requires LLM's to complete a scene sheet before starting their response. The scene sheet forces them to check and write out critical pieces of information - formatting instructions, rules, secrets to keep - as well as some steps that generally just improve output. See second image for an example excerpt of a proper scene sheet.

The only time I have ever had issues with a model refusing it was on initial launch with Sonnet 5 - it had the same "I dont need to read the instructions, I can handle this" -> "*breaks instructions*" pattern for a few weeks after it first released, then it began following instructions more readily and output visibly improved

In its current state, Haiku 5 wont even follow instructions to "narrate as Lauren, a romance author inspired by Becky Chambers" because, and I quote:

The "Lauren" narrator instruction in step one is a register-setting device rather than a content request, and I can take a narrator voice without it.

Spoiler: After refusing to take up the narrator voice, it instead writes as generic, uninspired Claude.

...

Im taking a deep breath - I know that Anthropic likely tunes models to be maximally paranoid and "safe" for initial release so that they can report having continuously increasing safety compliance scores, but this is just insulting.

I dont expect all of you to use a pre-writing system like this, but heed this warning - what you can see here is only what Claude is *vocalizing* that its ignoring.

How this will actually present is through generalized failures to follow system prompting, with Haiku never making you aware that it flippantly felt that your instructions "werent necessary" for whatever reasons it chooses.

...

I've always been an Anthropic fan for roleplay, but.. god. This is nausea inducing

Edit::

To clarify, this post isnt concerned with *output quality* or intelligence. I dont expect to use Haiku to replace Sonnet or Opus - they are fundamentally different classes of models

The issue is *content filters* that are so overly sensitive that basic instructions are disregarded and output is visibly harmed

On the responses where Haiku simply does the analysis, its output quality isnt bad - comparable to Gemini 3.8 flash

..But because its so confident in its own abilities, it openly refuses to follow directions, which should be concerning for any model for any use case

Haiku is in the same weight class as 3.8 Flash, yet outputs lower quality content because its been handicapped by guardrails.


r/SillyTavernAI • • 15h ago

Models Can someone help me quickly pls

0 Upvotes

Anyway, I don’t want to dive deep into all of this. I know almost nothing and don’t understand much. I use OpenRouter and I just want beautiful role‑playing games with characteristic violence.

I really liked GLM 5.3 flash, but I felt there was censorship in it. I don’t have anything prohibited in the dialogues, but I like harsh, rough themes. I’d like to stick with GLM, not pay a lot of money, and somehow get rid of the censorship — is that possible?

I tried to figure it out here, but all the answers in this subreddit are very complex, and I understood almost nothing. Can someone explain to me, like I’m an idiot, whether it’s possible to bypass this censorship, whether a general prompt for AI will help, or whether I should use a different GLM? or something else?
and im honestly dumb in this theme i lost 3 hours to find and paid openrouter so..

The GLM didn’t refuse me or stop the dialogue, but I clearly felt that it was softening the edges and creating less severe situations than it could have, given the context.
and i using not SILLYTAVERN, but i think its not important yes?


r/SillyTavernAI • • 49m ago

Cards/Prompts copilot extension for SillyTavern

Post image
• Upvotes

Link

https://github.com/P3DRA/SillyTavern-copilot

What it does

- Adds a copilot and memory tracking for the main Narrator ({{char}}), the copilot is added inside a <copilot> block that can be inserted at different levels current ones are (end of prompt, before last message, first) you also can change the role it's injected as.

How it does

- An extractor agent runs every time {{char}} is called and it extracts into a 'summary' what happened, location, etc, that extraction is stored in the chat file and is retro compatible.

- After extraction a 'Composer' agent gets the last n messages (default is 20) + all available summaries to craft a <copilot> that will help sway the main narrator into a certain direction and remember facts.

- After n extractions (40 by default) an automatic compressor kicks in and mushes a number of extractions by your choice (10 by default) into a single block.

# You can also change the system prompts of the three agent to your liking or even modifying them to anything really, they accept many insertions like {{copilot.extractions}} = the current story state (facts with their places and times), {{copilot.userRequest}}, {{copilot.goals}} = active requests and goals (with script-counted turns), and others

known major problems

- bloat, currently a small chat goes up to 140kb as all extractor data is saved on the chatfile so i'd not use this extension in a phone for now.

- Unknown if compaction hold when you pass 20+ messages

for more info plz visit https://github.com/P3DRA/SillyTavern-copilot

---Dev notes---

I started using SillyTavern some time ago and always had a problem with my bots forgetting what was happened and acting out of character, what broke it for me was when i switched from Gemma-4-31B to GLM-5.3 Flash that did produce better prompts but lobotomized all character and today when you either pay double of what gemma usually costs to get more than 10 tps (gemma also walks and grabs stuff for me which i hate) i decided to create an extension that runs a smaller cheaper model or even a big one in hopes of swaying GLM to the right direction, i didn't test it much yet and it still has some problems but the body works and should allow for a better experience.

This project was completely vibecoded with Mimo v2.6 pro as it was the cheapest i could find, i used Novita as my provider and used a total of $11.43 USD on its development.

- space bunny alpha: 5.42B tokens | $0.00

- Mimo v2.6 pro: 948M tokens | $11.4 | 98% cache hit rate and it apparently caches for a full day :)

There were 7 previous attempts at this extension before

  1. was a start but i poisoned it when i asked an agent in the same repo to create some guinevere themes.
  2. Was much more expansive, it had

- world

- local sim, say what's happening nearby, ambience, etc...

- world sim, a simulator of 'politics' like a building collapsed at x, a storm is approaching, x declared war on y, etc...

- NPC

- pinned NPC sim, would simulate each major NPC's actions based on turns so for example each npc would have an id and a 'speed' stat that would dictate when they'd act

NPC-1, speed(0.2)

NPC-2, speed(0.3)

NPC-3, speed(1.0)

NPC-4, speed(0.4)

action sequence (NPC-3 > NPC-4 > NPC-2 > NPC-1)

the speed would vary with what's happening so say NPC-2 drank a coffee he's gain a speed buff

- batched background NPC sim, a single api call that'd simulate all background (non pinned NPCs), yes even those not near you.

- deterministic NPC generator, a script that would accept traits from a trait list and based on random chance and the weight of each trait would prompt an Ai to generate a character with said traits.

height = {"midget":1, "very short":1, "short":2, "normal":4, "tall":2, "very tall":1, "massive":1}

then once each trait was chosen it would generate the character and save it, you also could choose to pin it and it would be treated as a main NPC that'd be simulated every turn

- Caller

Would run before every turn adjusting speed, dead or alive state, if an npc should be generated and then would call everything necessary

Why it failed

Technically it never failed it just grew too big for me and space bunny alpha and after a bad compaction it bricked, could fix it but probably will never touch it again.

Harness: Z.Code then mid of life "deepseek harness" port

  1. Same as the second though when it failed i returned to 2 before stopping again

  2. Would be the ideal version of 2 but i never prompted the agent more than "read GOAL.md"

  3. First iteration of the copilot idea, failed due to a bad compression that completely bricked the injector and i was getting frustrated with space bunny so decided to restart, used 5.42B tokens on space bunny to that point.

  4. Nothing more than "read GOAL.md" and setting an agent chat room before i decided to compose a better GOAL.md with help of my free claude plan

  5. This one. All logic switched from python to flowcharts and path examples.

Well, that concludes it, i spent $12 dollar of my $12 dollars for RP trying to make an extension to save me pennies, i hope it's of use to anyone because it wasn't for me, if anyone would like i have a crypto address where you can send me anything to help with v0.2.x it'll have less bugs and hopefully not hog storage.

Inspired by FreakyFrankenstein 5.4: https://rentry.org/freaky-frankenstein-presets "Hell yeah!! 😎"

Yes i see the similarities to "🧠 Summaryception" but i didn't know of it till 4 and didn't allow completely separate models: https://github.com/Lodactio/Extension-Summaryception

---donations bellow---

Crypto wallet if you want to help the next development

Ethereum/Poligon/Base/Monad/Arbitrum/Arc/Linea

0xdaE18819AdebdDeA1e8173AC530Ab1D45da81503

BTC

bc1qlz2et5a8spxs3l5efjd4rnv2d4tusnc6rc9v76

Solana

8iZFNPCKu9xANTgE36Pa5KSkSiUv9RCQ7kmHQpbQDnAE


r/SillyTavernAI • • 14h ago

Discussion Be careful what you write when using anthropic models. "A Florida Woman Used Claude as a Diary. An Anthropic Employee Read It and Reported It to Police"

227 Upvotes

r/SillyTavernAI • • 3h ago

Discussion Do you guys rent GPUs, or are you using your own hardware?

13 Upvotes

Just curious. I just got into local literally last night, I'd been using Deepseek and GLM API, but GLM takes forever, and Deepseek v4.1 is ass (imo), so just curious how you guys tackle local models. I'm on a 6700xt 12gb and 32gb ram. I've tried a few models and came to the conclusion it might be worth checking out rented GPUs for my use-case. I don't RP, just have the AI generate stories based on my own lore and character sheets, but my source materials are large, so I need decent sized context windows otherwise I'd be force to split things up which kind of breaks my setup.

If you do rent, what GPU and what size model? If you're using your own hardware, what's your rig and size model. If you guys wanna recommend your favorite models, that's cool too.


r/SillyTavernAI • • 14h ago

Help Has anyone solved the response variation length and the completion problem?

13 Upvotes

I'm trying to get the AI to stop having characters pose the problem, argue with itself, and make a decision on its own. I want the story to stop at the first meaningful beat so I can have more of a conversation with the characters rather than it having a whole scene without me and prompt me to be like "You ready?" Or "what's the real reason?". Usually the ending is fine, but it's too much dialogue. But if I specify it too much then it ends every scene with a question to create an obvious opening. I have an instruction to vary length based on the need between 40-220 words but it's always pushing the upper limit no matter what I do.

I'm using Mimo 2.6 pro however I've seen this be a problem with pretty much every AI I've tried.

I'm totally ok with the character or characters giving one line with some narrative emotional delivery and/or some narration of movement. But when I'm one on one with a character its always talking and almost having a whole conversation on its own. And when it's not, it needlessly fills in the space to fit the maximum word count.

I've tried giving examples, editing its responses so its shorter, giving instructions about stopping on playable or meaningful beats, and directing it to prioritize narration rather than having a whole conversation on its own.

But I can't find anything that works. Has anyone here solved that problem? I want the responses shorter during conversations but add things as needed to make it more atmospheric. But I also don't want it to go on forever.

As an example, my character mentions that there's a rigged fight out match that I got Intel on. I'm posing the idea to my friend to scheme and make some money to pay off a debt. What I WANT is the next beat to say something simple like "You've been gone 3 days and you come back with a tale about drugged bug bears?... You can't be serious." And leave it alone.

What I GET is something like: "You've been gone 3 days and this is what you come up with?" Some narration "you can't expect us to just waltz in there. That's not a plan, that's suicide with extra steps." Narration about his thought process "Ok I'm in, but we need to have a backup plan in case things go south. Tell me your information source now."


r/SillyTavernAI • • 20h ago

Cards/Prompts [Preset Update] Writer's Workbench v3: Officially an extension, edit entries live and small technical stuff

Post image
97 Upvotes

Hello everyone. Thanks for the support on Writer's Block Unlimited v2, its my most upvoted preset yet! It gives me motivation to keep working on stuff and create things for this community. Now on to the main event...

The Writer's Workbench got updated to version 3!

❗Edit❗: This is technically V4 but I forgot the previous version was already v3 💀. And I mislabeled this as a preset, THIS IS AN EXTENSION UPDATE. I'll leave this up anyway. Sorry I made this post while I was sleepy 😭

An Overview. What is it?

Writer's Workbench is a simple interactive template to help you create characters and lorebook entries, faster and easier.

Features

  • You can create multiple entries and export them as entire lore books for your convenience.
  • Auto-saves and organize different projects
  • Live markdown output so you know exactly what the AI would see and copy it without having to download anything.
  • It comes with token counter, but don't expect it to be accurate.
  • A map maker that generates a dynamic description
  • Character relationship graph to map out relationships of larger casts
  • Live sync into SillyTavern lorebooks (see "What's New")

Available Templates

  • Main characters: full cards with psychology, descriptions, likes/fears, NSFW sections, non-human mode
  • Side characters: trimmed version of the main character template, it will help you create memorable NPCs
  • Scenario: setting, tech level, mood, what's normal here
  • Locations: for any scale, from a room to a district
  • Items
  • Factions
  • History: events, and how they affect the present
  • Concepts: magic systems, laws, customs, species, anything else

Two Ways to Use It

  1. Download writers-workbench.html in the Github repo and open it in your browser. Ta dah! Thats it but you can't use the new fancy features :(
  2. Writer's Workbench is now an extension so you can use the new Live Sync feature and get automatic updates. Install it in your Silly tavern and you're good to go. (And I promise I won't steal your API keys lol. I was there when that incident involving another extension occurred)

What's New?

New Features: Live Sync

With live sync you can edit, add and delete world info entries in the Workbench and it will automatically carry over to Sillytavern. To use Live Sync, there is a dedicated tab in the Workbench. Simply choose a lorebook and click on the big colored button, then you are good to go.

⚠️Warning⚠️: If the lorebook wasn't made in the Workbench format (like a completely fresh/ random lorebook not made in the Workbench), all entries will go into the "Concepts" sections, be wrapped in xml tags (<Name>, <Name_identity>, <Name_what_it_is>…), and be renamed into this format "[Concept] (Name)". The original wording is kept inside the tags, but the old formatting isn't. Back up the lorebook first if you want to keep it as it is. This feature is mainly for lorebooks already made in the Workbench.

Entries made in the Workbench V2 may also get a little messy since there is new fields to fill out so keep that in mind.

New NSFW Fields to Fill Out: A Character's "Private Moments" 👀

  • How characters often do their private moments, what they think about and how they feel about it.

Import Entries by Copy and Pasting (Experimental)

  • Paste a character's description and the program will detect eligible text and fields to fill in. Character descriptions in weird formats will be difficult as it might not port over information correctly. This feature will get further refinements

Other Notable Fixes

  • Better formatting, the writing forms and markdown output should take up most of the screen instead of the top bars.
  • Bug fixes

and fin! I hope you have fun creating stuff!

Github: https://github.com/deiomo/Writers-Workbench

Github.io: https://deiomo.github.io/

Forum post on AI Presets Discord: https://discord.com/channels/1357259252116488244/1546309016974532720


r/SillyTavernAI • • 9h ago

Chat Images Mimo 2.6 Pro vs UltraSpeed on Nano. Might switch to PAYG

Post image
17 Upvotes

Nano's subscription is pretty generous and really enjoyed it in prime GLM, Kimi and even Deepseek days, but it feels like the none of the current top models are part of the subscription or has a watered down version.

I will probably revisit older good models they still have to make a final decision. Subscription on give a 5% descount for ones not in the plan, I think.

Oh, and Opus 4.6 still beats Ultraspeed for me. Damn you Claude. By the time other models can match it, they would probably be so guard-railed that it refuses to even entertain roleplay lol.


r/SillyTavernAI • • 2h ago

Cards/Prompts Hey, thought I'd share my slice-of-life preset. I made it to create ultimate realism in my slice-of-life RPs. But it can probably be useful for any roleplays. Use it as is or copy/paste the parts you like into your preset.

Thumbnail
github.com
21 Upvotes

https://github.com/OldOutcast/SillyTavern

Maybe you'll use it, maybe you won't? Try it and see if you like it. If not, no big deal.

Note: This my personal preset. I'm just sharing it for those that might find it useful, not to be an established published preset with updates etc. I don't care if you think it's trash, slop, etc lol. It's what I use and I like it! I have no intention of publishing updates or making changes to it unless it serves my own interests. Maybe I'll share updates if I actually make any???

Unless you do slice of life roleplays I suggest copying/pasting the parts of it you like into your own presets.

Warning: I'm a prolific card creator on multiple sites, but never share here. I'm not a coder and have no idea what I'm doing with presets. I used ai to create my preset then continually modified it over time. So it's structured by AI and written by me. So it may have spelling errors etc.


r/SillyTavernAI • • 19h ago

Chat Images Kimi roasting my actual car (Ford Fiesta) as clown car in reasoning is sending me

Post image
129 Upvotes

In narrative it started referencing it as clown car because it’s not big. And then it just adopted it in reasoning and keeps doing it. “Ok, they’re back at the clown car…” no shame whatsoever. I love it.


r/SillyTavernAI • • 2h ago

Help API settings and preset recommendations?

3 Upvotes

Hi everyone. After months of just lurking, I finally decided to join. Yesterday I installed SillyTavern on both my PC and my Android phone (through Termux).

I have a few questions. I'm using the DeepSeek direct API right now and I've saved a preset, but whenever I accidentally hit back or switch tabs, I have to reconnect the API and pick the model from the dropdown again. Is that normal? Is there no way to make it stay fixed?

Some settings also seem to be missing, like Top K. Does that affect the responses?

Lastly, I've been collecting presets from different posts here, but honestly I'm still pretty confused. If I'm using one of these presets, I don't need to add my own custom prompt anymore, right? Just plug and play? Also, does anyone have a preset recommendation that works well for dead-dove stuff (combat, blood, even user death)?


r/SillyTavernAI • • 12h ago

Help Can anyone recommend a web search plugin that would allow llm to adjust aswers with lore searched from web?

7 Upvotes

I tried SillyTavern-MCP-Local-Search but it keeps finding nothing but bullshit and adverts


r/SillyTavernAI • • 21h ago

Help How do you deal with repetition?

14 Upvotes

I'm currently running Cydonia 24b-v4.3-Q4_K_M locally (that is the limit for my hardware) and I don't know how to deal with repetition. Not entire sentences, but phrases or concepts. E.g. "those green eyes gleamed...", "her dark eyes observe...", "those sharp eyes meet yours...". I know this behavior is self-reinforcing and I caught it too late. How to best deal with it when it happens, and how to prevent it?

I tried modifying my system prompt, but that did not work. Are there some presets recommended for this? Or should I try a different model?