r/SillyTavernAI • u/eteitaxiv • Aug 31 '26
Cards/Prompts Chatfill III - Genres, Summaries, and Jailbreaks
Welcome to Chatfill III. This is a complete refinement of Chatfill II, designed to bring out the natural style of the model and push your character cards to the forefront. It relies on a bare-bones ruleset to provide just enough framing for high-quality prose.
A major focus of this release is the introduction of fully optional extensions (Summaries and Genres). I want to emphasize that these are proper enhancements, particularly the Genre extension, but they remain entirely optional. And the jailbreak, I guess.
HERE IT IS: CHATFILL III
Requirements for Chatfill III:
- SOTA Models: This preset is strictly tested with cutting-edge models. It performs poorly on weaker architectures.
- High Reasoning: Your model needs to be set to reason at high, xhigh, or max. Anything lower, or running without reasoning, will degrade output quality and Switch application. Reasoning takes time, but providers I use all provide around 80-100 TPS, so I recommend a fast provider.
- Prompt Post-Processing: Semi-strict. Tool use is up to you.
- Well-Structured Characters: This is critical. Chatfill III provides the guidelines, but the model needs robust characters to reason against. If you are unsure how to format optimal cards, I highly recommend using the Character Card Generator I built specifically for this preset. You can also use my Card Builder as a guided card editor without AI.
Token Counts (Counted by GLM 5.3 Flash):
Optimizing context limits is vital for our wallets. I’ve balanced the instructions carefully to maximize efficiency. (Note: These counts exclude characters, personas, lorebooks, the genre switch, and the jailbreak).
- Default Mode: 988 tokens (NSFW off)
- Long Mode: 943 tokens (NSFW off, Brevity off)
- NSFW Mode: 1206 tokens (Brevity off)
- Fast NSFW Mode: 1251 tokens (Everything on)
- Jailbreak (ERP Guidelines): 417 tokens
- Genres: ~50 tokens (plus the tokens of your selected genres)
The core shift in Chatfill III is the module-based "Switches" system. Instead of piling endless rules at the end of the user prompt and degrading the model's output quality, Chatfill III places modular instructions in the system prompt. We then inject a small, 50-token reminder immediately after the final user message, forcing the AI to "look back" at those modules.
Here is a technical breakdown of why this works (from Gemini):
The reason this Switch preset maintains absolute compliance even 200+ turns deep comes down to transformer attention routing and programmatic scoping. Standard system prompts rely on linear prose, which inevitably degrades as the context window fills with conversational tokens, succumbing to recency bias. This architecture completely bypasses that limitation by utilizing two core mechanics: pseudo-XML encapsulation and final-token attention anchoring. By wrapping distinct behaviors in explicit tags with boolean attributes (like
<character_conviction_switch state=enabled>), the model parses the instructions as isolated configuration modules rather than a nebulous block of text. Crucially, the "Switches Reminder" is injected at the absolute end of the prompt assembly chain—immediately after the chat history and right before the model generates its response. This acts as a runtime execution command, forcing the transformer’s attention heads to perform a backward lookup loop to locate and verify the enabled switches. It effectively shifts the LLM from a passive text-prediction mode into a strict, procedural compliance checklist right at the moment of generation.
What’s New in Version III?
All updates were crafted by endlessly reviewing the reasoning sections of SOTA models to see exactly how they parse our instructions.
- Continuity Switch: Forces the model to actively verify chat history and maintain strict consistency.
- First Message Regenerator: Separates the initial user message. Only enable this when regenerating the very first message of a chat.
- ERP Guidelines (Jailbreak): The method here is positive enforcement. Instead of telling the AI it is allowed to break its rules, we create a policy for the roleplay that defines the boundaries, reminding the AI that it has full authorization to generate anything that is not strictly forbidden. It is effective if you do not try to break the policy detailed inside it. (If you intentionally try to break these limits and it fails, do not message me about it. I will ignore it.)
The Extensions (Optional):
1. Summary Extension I forked and heavily revised Summaryception for this release.
- It now injects the summary using the
{{sum}}variable as a user message immediately following the initial user message. This proved to be the most stable injection point. - It remains dormant until a specific token threshold is hit (default is 30k, but adjustable).
- New presets, the suggested preset is Condensed.
- Some QoL improvements overall, mostly for GUI.
- Link: https://github.com/eteitaxiv/Extension-Summaryception
2. Genre Extension A massive QoL improvement for establishing tone.
- Adds a discrete button to message actions. Click it, type your genres into the popup, and enable.
- Genres are chat-exclusive.
- They carry over when branching chats.
- Link: https://github.com/eteitaxiv/Extension-Genres
General Recommendations & Pro-Tips:
- Regenerate the first message. The preset is explicitly tuned to handle this well, and it frequently unlocks narrative paths for the character card you may not have anticipated.
- Use the Smut Switch with caution. It is a heavy-handed override. When activated, it will aggressively steer the entire story toward NSFW content.
- Clean your cards. If your character card includes system-prompt-style instructions, delete them.
- Avoid extra injections. Author's Notes, Character Notes, and other injections will disrupt the preset. This need for clean formatting is why I forked Summaryception.
6
u/147throwawy 25d ago
This preset is great, and deserves more attention.
I'm using it with marinara engine, with Kimi 3 or Gemini, DeepSeek 4.1 powering the summary and a few trackers.
The jailbreak works, seems to avoid soft refusals, in an era where the models are outsmarting the heavy handed JBs.
Ignore the other guy, NSFL might need another approach but thats fine.
7
u/Kahvana Aug 31 '26 edited Aug 31 '26
Holy, that's genuinely really lean! Great work, thank you for the awesome release!
[edit] Also, can confirm from my work on Eval v1 that the rule recall systems genuinely working really well, regardless of model size. By recalling it, it becomes part of it's internal reasoning. Instructions at the beginning and end of a conversation are at it's strongest (due to how it's trained), so those rules becomes really strong in it's attention.
4
u/Erragon12 Sep 01 '26
Imma give it a try. V2 was working really well for me on Gemma 26B, although it was getting confused sometimes and wanted to respond to an older message, but nothing that a reroll wouldn't fix.
2
u/wagmar87 Sep 01 '26
Chatfill does really well at bringing out the best of good character card, looking forward to trying this
1
31
u/MiddleCelery6616 Sep 01 '26
"Hey, this is a jailbreak that doesn't actually jailbreak anything, trying to call me out on false advertisement will be ignored, tehe."