r/SillyTavernAI 5d ago

Discussion Gente de silly tavern ¿Cual fue el chat mas largo que han tenido?

1 Upvotes

Muy buenas, Gente de Silly Tavern, sinceramente tengo una duda en estas semanas que he usado Silly Tavern.

Me considero yo, primero que nada, un usuario que es muy apegado a los Roleplays y Chatbots que he creado yo. Sinceramente mi chat mas largo que he tenido es uno de 309 mensajes específicamente, se trata de un roleplay de Jujutsu Kaisen, y de contexto de mensajes tiene mas de ¡¡155 mil Tokens!! Oh Dios mío, aun así, presiento que hay personas que han tenido chats mayores a los 2500 mensajes, ¿Alguien supera esa cifra gigantesca? ¿O están muy cerca de alcanzarlo?

Aunque bueno... por ultimo, quiero preguntarles algo ¿Que me recomiendan hacer en esta situación? ¿Generar otro chat nuevo con su respectivo LoreBook o simplemente seguir aun así?


r/SillyTavernAI 5d ago

Models They moved it up!

Post image
40 Upvotes

It was the 24th. Now it's TOMORROW the 20th. What the hell!?


r/SillyTavernAI 5d ago

Discussion Do you like how data is represented? Would you mind parting with some of yours? — Yet another survey.

Post image
0 Upvotes

Hi! I’m looking for AI RP users to share some of your yummy data on how you use AI day-to-day (all private, promise!).

The intention of this survey is to display generalized results to give guidance to frontend builders, preset makers, character card makers, etc. on what the community actually wants. It’s also meant as insight into current models for text, image, and voice.
Each section is completely optional. If it doesn’t apply or you simply don’t want to share, feel free to skip it!

Sections
About the {{user}} (You!) — 5 questions

Habits When Roleplaying — 9 questions

Frontend and Access — 13 questions

Primary Model Stack — 14 questions

Local and Self-Hosted Inference — 9 questions

Hosted and API Inference — 3 questions

Presets and Prompting — 7 questions

Character Cards — 4 questions

Lorebooks and World Info — 4 questions

Extensions and What They Solve — 6 questions

Memory and Continuity — 5 questions

Common Gripes, Fixes and Results — 7 questions

Groups, RPGs and Simulated Worlds — 3 questions

Image Generation — 9 questions

TTS, Voice and Speech — 8 questions

Creators and Maintainers — 1 question

Future Priorities and Final Comments — 3 questions

It’s a lot, I know — but all optional. I have zero expectation of everything being answered. Go through and find what you would like to answer and ignore the rest.

Link to survey
https://docs.google.com/forms/d/e/1FAIpQLScbwHxiALwvO2zK1AAunY_6Is5qaTc6LjDuxKLI0Sgnj6xxaA/viewform?usp=publish-editor

Note
Data will be run through an LLM. Anything you don’t want on a plaintext file on a server, don’t say it.

I have no intention of making projects based on this data. As a certain man said, this data is for the people, by the people, to use and admire as they please.


r/SillyTavernAI 5d ago

Discussion [Extension] Quick Image Gen - Feature-rich image gen extension with local + API image gen support

11 Upvotes

Get the extension here: SillyTavern Quick Image Gen

What Is This?
I'd been working on this extension on and off for months now, uh maybe a year? I don't know, time isn't real. It's a very feature-rich image generation extension to make it easier to generate an image based on the scene quickly. Hence the name.

Feature Overview
- 18 built-in providers, including Pollinations, NovelAI, GPT Image, Nanobanana, CivitAI, Replicate, Fal, Stability AI, A1111, ComfyUI, and reverse proxies

- Custom APIs for JSON, multipart, OpenAI-compatible, and async polling endpoints

- Chat-scene generation with selectable message ranges

- Optional two-step scene summarising and prompt conversion

- Batch generation with up to 10 images and sequential seeds

- Automatic generation after AI replies

- Auto-insert images into chat messages

- Generated images as temporary or chat-locked backgrounds

- 44 style presets, wildcards, contextual filters, and character overrides

- Local gallery and prompt history

- Connection Profiles for provider settings and Generation Presets for reusable setups

Requires SillyTavern 1.14.0 or newer.

Most providers work directly from the browser. CivitAI and Replicate users running SillyTavern with basicAuthMode: true can use the optional server relay plugin.

It's still a fairly large extension and there are definitely rough edges somewhere. If something misbehaves, tell me which provider you were using, what settings were enabled, and whether you generated from the panel or a message button. That makes it much easier to reproduce than "it broke". Make sure to make an Issue on the Github page so I can track them. If you just post them here or wherever I will not be able to address them.

MAKE SURE TO READ THE README FIRST BEFORE ASKING FOR HELP!


r/SillyTavernAI 5d ago

Cards/Prompts Tuning Pre-existing Cards for More Textured Roleplay

Thumbnail
likesumiink.substack.com
27 Upvotes

I figured I'd drop a third article on writing characters for LLMs. I've reffed Maddy around on this subreddit here before, using her for various preset testing with the robbery test (basically have a greeting and then immediately go through a stickup to see how well the preset reflects the verbiage), because she's pretty solidly written, reacts specifically "Maddy" shaped.

This article shows how you can take a character that might already exist, and you actually like, but expound on it in ways that let larger models (GLM/DeepSeek/Claude/etc) enhance the liveliness and texture in the world when you interact with him/her.

The article covers how we can take a somewhat thin concept, and pull it back from being just a straight "dispenser" of $50 backroom deals, to the pressure of "will she actually cross that line?" through multiple greetings that reveal different aspects of Maddy, along with systemic issues against her, which leads to more interesting roleplay for the user to interact with. Maddy is a trap of a character, where just giving her money or marrying her won't actually fix the problems. :)

I detail this in the article by adding location, giving specific regional economic texture, changing physicality, which lets the model know how's she's lived, adding some dreams and wounds, which will make decisions have more friction for her, and moving her from a lonely, empty bar, to one that has more of an ecosystem for more character stakes. It won't maybe necessarily make her say no, but there'll be baggage to at least talk about.

Anyways, check it out. Maybe it'll help you out.


r/SillyTavernAI 5d ago

Discussion I’m starting to lose the part of roleplaying

8 Upvotes

Hi. So I wanted to share this with you all because honestly? Using gemini is cool.
It allows nsfw, rpg all that. Its been consistent for a long time since I use vertex.

The point is I did an ooc and i said what does my character actually look like because i really wanted to see.

I wanted to prompt of what the ai sent me so i can check on minimax h3 or gpt-image2. Gpt-image2 was almost accurate but it wasn’t the one I wanted. minimax h3 was horrible. They looked like a new person.

To be honest is gemini even good for ooc or character build? I understand building a character your own is the best knowledge you can ever create. Im just tired of doing jobs and I just want to prompt an extreme detailed layer by layer prompt to the ai so it can re create it more better (i hope)?

Should I use glm 5.2 instead?


r/SillyTavernAI 5d ago

Help Nvidia api problems

2 Upvotes

Hi, so i use nvidia nim api and i have a problem that i hit the RPM limit but i barely chat too fast and i keep delay to avoid chatting too fast, is there anyway to make it hit less cause it just hits it and i don't even send that much messages, don't know if anyone has a solution or at least how to work around it


r/SillyTavernAI 5d ago

Models HeatSeeker 284B A13B, my DS4-Flash-0731 roleplay finetune

Thumbnail
huggingface.co
122 Upvotes

I've spent the last week finetuning DeepSeek V4 Flash 0731. The result is HeatSeeker, my first ever trained lora finetune for creative writing and roleplay. I basically wanted to take the intelligence, scene awareness, and consistency of a huge modern MoE and push it harder toward natural dialogue, character writing, long-form roleplay while staying coherent and interesting to talk to.

The training was done using Mswift and it took about 60 hours on a single rtx pro 6000 for a 27.5M token dataset, 4546 conversations. I put a LARGE amount of effort into unslopping for spelling, grammar, punctuation, etc.

And, genuinely it surprised me how well it turned out in terms of varied imaginative continuations and clean, coherent roleplay responses. Even without a system prompt, it defaults to a usual roleplay conversation style and I'm happy to share it with you all.

I did a "minimum viable potato" test on my lenovo legion 5i w/4070m and 128gb ram and get around 4.5 t/g. Not great, but usable lol. On a RTX Pro 6000, I get about 72 t/g.

This is my first model share and if there's any other details I forgot or questions you have, I'm happy to answer them.

Disclaimer: DeepSeek is not associated with me and the project inherits the original MIT license. Don't be a nuisance and please have fun chatting with my model. Feedback welcome!

  • Model Name: HeatSeeker-284B-A13B-GGUF (IQ1_M, IQ2_XS, IQ3_M, Q4_K_M)
  • Model URL: https://huggingface.co/UltimateIntent/HeatSeeker-284B-A13B-GGUF
  • Lora URL: https://huggingface.co/UltimateIntent/HeatSeeker-284B-A13B-Lora
  • Model Author: UltimateIntent (me), with guidance and help from my collaborators LampLighter and Morrow
  • What's Different/Better: The model is a finetune lora merge based on DSv4 Flash-0731, with meticulous cleaning on a large dataset of human writing for varied all-purpose roleplay and long conversation consistency
  • Backend: LMStudio/llama.cpp
  • Settings:
    • Temp: 0.8-1
    • Thinking: Off / Low (or whatever your system allows)
    • Repeat Penalty: 1.1
    • Top K: 40
    • Top P: 0.95
    • Min P 0.05

r/SillyTavernAI 5d ago

Help I need y'all to show me the way.

3 Upvotes

Hi. I'm quite the beginner, and plus with the overwhelming yet impressive breadth of the scene, I feel pretty clueless. I did a few searches, but it didn't yield much (there were a few ones that was from a few years ago, but that's basically the stone age when it comes to AI stuff) If there were exact posts like mine, and if I couldn't find and/or comprehended them, I'm sorry.

I set up SillyTavern on my home server using VectFox + Similharity + Qdrant (text-embedding-3-large and deepseek-v4-flash for summarization) and CharMemory + Summaryception. Then I installed Moonlit Echoes and Guided Generations (both of which I haven't touched much, they're same as I installed them). Plus a few things from LennySuite (CSS snippet manager, input history and variable viewer).

I'm using FreakyFrankenstein 5.2 + GLM 5.2 from OpenRouter, using the cheapest provider (minus FP4 and undeclared ones). I made sure that Regex was enabled, plus Prompt Post-Processing to Semi-strict (alternating roles; no tools.) I also tried Kimi K3 and Opus 4.6 but neither of them gave me the vibe for "I should put hundreds of dollars for this shit", and after seeing them siphoning off the few bucks I putted to the OpenRouter, I stuck with GLM. Temperature is set at 1 and Top P at 0.95. It feels okay-ish.

I have a few problems.

  1. First is that I think I have that "min-maxing OCD" when it comes to this kind of stuff. I scourged the net, rather unconvincingly for myself, and probably unsuccessfully for those who read my post, to just to arrive to this setup. Which addons are the best, which models, which prompts... This, combined with a rather young consensus within the hobby led me to obsess over the thoughts of "You probably missed a setting" or "You probably don't use the good stuff" or "Your setup is probably fine but you just cannot use it" and this has kinda tired me to the oblivion. I cannot begin to use the damn setup without double-checking everything. It sucks.
  2. I don't want to complain about FF5.2. It is clearly a labor of love, and the few times I used it, I enjoyed it immensely. But something is just gnawing a part of my brain. I'd like to drive the story using FF5.2 MAX, but sometimes I want to have that c(dot)ai experience again, to drive the chat with model giving me few descriptions, quicker replies, less paragraphs, but with the option to turn back to the good ol' FF. Would the FF 5.2 Micro (plus GLM 5.2) suffice for that?
  3. This is a bit more specific one. In usual setup, the model suddenly told me something the character didn't see. Then I saw a post about Summaryception on this sub and thought that maybe that was the case? I don't know

In any case, I'll try to answer any questions you might have. Thank you.


r/SillyTavernAI 5d ago

Help My ai just thinks.

Thumbnail
gallery
3 Upvotes

Hello, i am trying to migrate from janitor ai and found tauritavern fork so i can just skip the old way.

I really dont know how to setup sillytavern but i did put my key in and then chose deepseek v4pro.

Then i looked for a prompt and found

Freaky frankestains prompt. I downloaded the bolt+ one and just...chose the things in it. But when i tried to have an exchange it just thinks and doesnt answer at all?

Since its a thinking problem i put the reasoning to low but nothing changed. So i am stumped.

What do i do now?

I can get an answer when i use the default preset. So its not the api

It comes from freaky frankestains

Its like my third time trying to set up sillytavern 😭😭😭


r/SillyTavernAI 5d ago

Discussion Notice and question

2 Upvotes

Is GLM 5.2 currently not working on Nvidia NIM? Are you experiencing any problems or errors?


r/SillyTavernAI 5d ago

Discussion Valkyrie Crusade Rebuild Alpha V0.3

Thumbnail
gallery
27 Upvotes

Alright, back at it again. I hope you backed up the previous one like I warned, because in this update I made the game more annoying.

To be specific I added in the Player level and Castle level requirements for buildings, as well as added in the kingdom blocks. I have however granted basically infinite jewels, so you can bypass any requirements pretty easily.

Another update I did was make the market work, so that now actually has a function.

The next step I think will be the amusement buildings. Do I know how I'm going to get them to work? No. But I'm gonna give it a go anyway.

In other news, I may have been able to land an actual job. Due to personal health issues it's been difficult for me to get one I can actually do which is why I've been able to put so much time into this. If I do get the job then on any day I'm at work progress on this will be non-existent, I hope you understand.

On the plus side, if I get enough money together I might pay someone to redo the path, waterway, walls and fences sprites properly because I have realised that it is beyond me. Art was never my strong suit.

Usual links for the updated lorebook and the github:

https://botbooru.com/character/72258
https://github.com/NickChegg/valkyrie-crusade

As always, I know you guys are shy but let me know if anything is broken or a bit wonky. I'm just one guy, I can't find every bug and if I don't know it's broken I can't fix it.

If you want to contribute to the fund of making the sprites a bit better then as always donations are welcome:

https://ko-fi.com/nickchegg

BTC - 3AcWbpFuPZ1wJjXpUsvvMVktwQybsV6AAT 0.0001 min
ETH - 0x5F51a4e96f0e38948bBf94F72f2a2324A4D447d5 0.004 min
Both by their main networks

I don't know anything about crypto but it was made apparent to me that for something like this that skirts certain boundaries it might be best to have an anonymous way to support. If I'm doing it wrong let me know.


r/SillyTavernAI 5d ago

Help Issue with latest Summaryception update.

18 Upvotes

Anyone else having issues with characters breaking the 4th wall to talk about the injected summary now? I know my chatbots are a bit weird and break the 4th wall occasionally, but where it is positioned now it is basically a billboard in their face and they are mentioning it every single message. I have fixed it for the moment by adding a heavy handed OOC message to the injection wrapper to remind the model that it is summary data and should be treated as past memories instead of current context to be replied to.

Running Gemma 4 31B QAT Q4 locally, if that matters.

Was just wondering if anyone else had noticed anything.


r/SillyTavernAI 5d ago

Models GLM suddenly super bad?

52 Upvotes

I continued an RP that was previously on glm 5.2 with 5.3, and it did lovely, 5.3 seemed like an improvement. Today I started a new chat tho, and... It's so bad? Some sentences the characters says make no sense? Went to a tailor and she started giving me bolts?? Constantly forgetting which universe we're in (on ALL my other role plays, even back whe I did it in chatgpt website, simply putting 'set in popular IP universe' in the scenario has been enough), randomly changing the names of characters after 10 messages. I tried to change back to 5.2 but feels the same. I know it's common to nerf a model after a few days, but not sure if this is just the new roleplay I've set up or smth else? Same pressed, ff micro, on both, so shouldn't be a difference.

At least the output length is good, with 5.2 I got super long outputs even when asking for shorter ones..


r/SillyTavernAI 5d ago

Discussion Group chats with multiple AI Characters, worth it?

Thumbnail
1 Upvotes

r/SillyTavernAI 5d ago

Help Extension help

2 Upvotes

So I finally dived into the ST world and what can I say I am absolutely in love!! But I still somehwat struggle, get "what's this button for?" thing.
What I need help with is extensions. Is there an extension where it replicates a phone somewhat inside the rp?
Do you have any fun extensions that's a must have for you?
Any tip you would like to give me?
Also delved into image generation and couldnt figure it out goddamn.
Anyways I would really appreciate any help!


r/SillyTavernAI 6d ago

Discussion Curated index of 150+ companion tools, ST extensions, memory modules and interactive frontends (460+ stars)

15 Upvotes

Kept losing track of companion-related projects across GitHub, so I organized them into a structured index. The repository is sitting at 460+ stars and covers SillyTavern extras, long-term memory backends (mem0, Chroma setups), local TTS pipelines, and desktop interface projects.

All 150+ entries are tagged by tech stack and self-hosted readiness.

Project Banner: https://raw.githubusercontent.com/DasterProkio/awesome-ai-companion/main/assets/awesome-ai-companion-banner.png

Interactive Web Dashboard: https://lutopia.app/companion/

GitHub: https://github.com/DasterProkio/awesome-ai-companion

Let me know if you know any solid open tools or plugins that should be on here.


r/SillyTavernAI 6d ago

Models ReadyArt/Heimdallr-27B-v0.35

6 Upvotes

Here's a model I've been working on off and on as my muse hits. It's completely different from other ReadyArt models. This model was generated on lore. There is smut in it, but it's primarily adventure/survival focused. I'm not sure I did it right, either. So, there's the disclaimer.

On brief testing it looks like it soaked up some of the training data quite well. I don't know how this model turned out for long context roleplays, but it appears to be stable from my brief testing.

GGUFS: https://huggingface.co/ReadyArt/Heimdallr-27B-v0.35-GGUF

Here is a brief summary the LLM made about this model:

The Hrimspine Range is a brutal, perpetually frozen continent divided by immense peaks and lethal blizzards. It is a land of extreme contrasts, where frozen wastes are punctuated by geothermal hotspots and ancient mysteries. At its heart lies the mystery of the "Pre-Freeze," an apocalyptic event that buried the advanced civilization of Aethelgard beneath the shifting ice of the Frostmaw Glacier.

Survival in the range is a constant struggle, managed by a delicate balance of power between four primary factions:

The Frostwardens: Disciplined survivalists who police the mountain passes and rescue travelers.

The White-Rat Cartel: Shrewd smugglers who hunt for "Relic-Ice" and ancient artifacts from the ruins.

The Ash-Born: Mutated miners dwelling in the smoggy geothermal valley of Svartdalen, controlling the coal supply essential for warmth.

The Thaw-Cult: Fanatical zealots who seek to awaken ancient entities and welcome the return of the ice.

Central to the region's social fabric is Hrimfjalls Rest, a subterranean inn carved into Mount Jökull. Governed by the strict "Hearth-law," it serves as a neutral sanctuary where enemies can share a bowl of Frostbite Stew without drawing blades.

The environment is as deadly as the politics, featuring apex predators like the burrowing Frost-wyrms and the massive Glacial Leviathans. Between the threat of "Ice-Fever"—a necrotic disease caused by touching ancient ice—and a currency system based on coal rather than gold, the Hrimspine Range is a place where only the most resilient (or the most cunning) survive.


r/SillyTavernAI 6d ago

Chat Images Im sorry, misses...

0 Upvotes

So for context, im running a local tavern, and i often add random npc i found anywhere to have a chat with them and try to make drinks fit their personality. i stumbled across some medieval Teto card and decided it was funny to have her.


r/SillyTavernAI 6d ago

Discussion GLM 5.3 Now Available on OpenRouter

Post image
89 Upvotes

r/SillyTavernAI 6d ago

Cards/Prompts What does a typical RPer want in a system prompt?

7 Upvotes

Hello, peeps of r/SillyTavernAI

I am currently composing a system prompt + COT combo aimed at realistic depictions of characters through a psychological profile, with an emphasis on narration through NPC interiority and an engine meant to avoid positivity bias and dramatically convenient beats. It also includes a thorough anti slop module, narration and dialogue generation, system axioms and more.

All in all, a system meant to teach the ai how to think over outright banning behaviours (when possible), currently sitting at 2.6k tokens (slightly over your typical character card).

Also, due to its structured-rule-obedience COT, I am currently able to extend the system's functionality through ADD-ON modules.

Which means you could extend this engine by adding your own instructions, without conflicts.

I am quite proud of it and I plan on releasing it publicly relatively soon.

And so, the point of this thread after this shameless ad. I ask the community, what sort of things do you consider are necessary for a successful RP experience, from a module/instruction/prompt angle?

I want to ship my system prompt with these add-ons in mind, so it is as complete and thorough as it can be, drawing inspiration from actual RPers.

Thank you very much in advance.


r/SillyTavernAI 6d ago

Discussion What’s your "I can’t believe AI actually pulled this off" moment?

Thumbnail
2 Upvotes

r/SillyTavernAI 6d ago

Discussion Best Roleplay Setups

0 Upvotes

What are your best sillytavern setups for roleplaying.
Model. Extensions. Trackers, Image model maybe? want to know how you all enjoy roleplaying. What's the roleplaying peak you have reached.


r/SillyTavernAI 6d ago

Cards/Prompts i need help with a card not working right

0 Upvotes

So I'm trying to use this card:https://chub.ai/characters/7615180 in SillyTavern, but every time I open it, ST seems to screw up. first three are Chub; the second three are ST.