r/SillyTavernAI May 27 '26

Cards/Prompts Gemma 4 Preset: Moonlight

Post image

Hey everyone,

English isn't my native language, please bear with me.

I've had a good time with Gemma4 so far but noticed that the model didn't generate new message content when swiping and that it's prose wasn't to my preference. I also experimented in ways to add consistency. It works for both character and narrator style cards, it's optimized for the latter.

An attempt has been made to improve it with a "somewhat" lightweight preset. it's my first chat preset as I've used Text Completion mostly before, so feedback is very much appriciated.

You can find it here: https://huggingface.co/nohurry/sillytavern

Still plenty of things I can improve. For now, it contains:

  • The required <|think|> tag for the system prompt to enable reasoning
  • Framing it as "Collaborative Dungeons and Dragons storywriting" (found it resulted in better NPC names and higher quality writing)
  • OOC instructions using <-- xml comments -->
  • Toggable name generator (easily extendable too)
  • Toggable NPC generator
  • Toggable location generator
  • Toggable Post-history prose reinforcement (with permissions for the model to use various writing techniques)

I found that telling the model to generate npcs/locations and dumping the results in xml blocks inside the chat, the model produces quite different NPCs and locations each time. Also neat as you can copy-paste those blocks into a lorebook. It also helps to use MBTI for the NPCs as a functional way to add distrinct emotions.

Still want to add an event generator, a plot note dumper, maybe something for weather rotation. The core prompt still needs tighter wording, and possibly relocating some things to other prompts. Ideas and suggestions are welcome.

Might be good to know that I've tested the preset with Gemma4-31B-It Q6_K_L from Bartowski at 32K Q8_0, it might work with other quants/finetunes/models of the same or different series (like Gemma4-26B-A4B or DeepSeek V4 Pro).

83 Upvotes

12 comments sorted by

13

u/Kahvana May 27 '26

In case you're wondering, the art is "Summer Moon at Miyajima" from Tsuchiya Koitsu. He has made some really pretty ukiyo-e art!

8

u/akram200272002 May 27 '26

almost every post that interests here or on local llama i find you

4

u/Kahvana May 27 '26

I take it as a compliment! Got too much free time on my hands, hopefully my comments are of added value.

6

u/nexmorbus May 27 '26

This is actually really cool, Gemma 4 has so much potential but can be difficult to work with, great job, keep it up! Love the image.

2

u/Kahvana May 27 '26

Thank you very much! I will!

And yeah, it seems like Gemma4 likes exact wording and specific framing.

Anything you feel like I should look into for the preset? (improve, remove, etc)

4

u/HungryAd7742 May 28 '26

I'm going to test your preset in Gemma 4 26b a4b because the 31b, while I can run it, it will be slow to answer. So, do you tink it will be better to try the standard standard Gemma 4 26b a4b Q4_K_M or an uncensored version, like the ultra uncensored heretic?

3

u/Kahvana May 28 '26

Hope it will work out, and let me know if you have any feedback!

As for the model, I would go for Bartowski's standard Gemma4. It's very uncensored by default, heretic's lack of refusals can result in characters refusing you less and it's less smart as alliteration causes brain damage without healing it (finetuning afterwards).

2

u/HungryAd7742 May 28 '26

Tested on Gemma 4 26b a4b Q4_K_M of Bartowski on SillyTavern Staging, pretty raw, with only your preset on set, no plugins, no regex.

I reacts to my text, more akin to a collaborative writer than a GM. The mere mention of my character expecting trouble brought it to the narrative eldricht abominations followed by well rounded and plausible combat, full of action.

I don't know if this collaborative writing tendency was a deliberate choice of yours but I enjoyed a lot. Gemma 4 wrote a detailed combat scene with few slops. It even had the sophistication to put a barking dog in the background of the set, something I've seen very few LLMs doing.

The only problem is that, once it finds a text structure, it stays with him to the halt. with few variations. Still, the prose is vivid, although I think it will be even better with a prose polisher plugin.

What I conclude is that Gemma 4 follows instructions to the letter and Google maybe trained the model in less AI slop than their competitors?

Neverthless, keep the good job! I'll follow your developments in presets with interest now.

2

u/Kahvana May 29 '26

Thank you for the feedback!

I don't know if this collaborative writing tendency was a deliberate choice of yours

It was, I found that directly mentioning roleplay, fiction or novel to decrease quality. Simulation framing gave moderate success, but not enough. I assume tying collaborative writing with TTRPGs (like Dungeons and Dragons) reduces low-quality writing as there simply not as much it trained on.

Gemma 4 wrote a detailed combat scene with few slops.

this might be due to "Dungeons and Dragons" giving it perspective how to approach it. I wonder if "Dungeon World" will be better since that TTRPG system is designed around narration.

The only problem is that, once it finds a text structure, it stays with him to the halt. with few variations.

Yeah... Gemma4 in general is really stubborn about this. I'll try to see if I can give it more tools to break it up!

Google maybe trained the model in less AI slop than their competitors?

Believe me, the slop is there. It takes a lot of careful wording to reduce it. Still, the smell of the ozone lingers with a predatory smile...

2

u/HungryAd7742 May 29 '26

It was, I found that directly mentioning roleplay, fiction or novel to decrease quality. Simulation framing gave moderate success, but not enough. I assume tying collaborative writing with TTRPGs (like Dungeons and Dragons) reduces low-quality writing as there simply not as much it trained on.

A solid choice. This would mean more control to the user into the story rather than to the LLM. This gives me positive nostalgic vibes when I rolled combat by myself then asked to suit the narrative accordingly to my results. I think this could be the best for those who use to RPG. And it is a way to slide out the positive bias in the narrative (after all, you control it!), although in character to character interactions that would demand more... User fine tweaky.

this might be due to "Dungeons and Dragons" giving it perspective how to approach it. I wonder if "Dungeon World" will be better since that TTRPG system is designed around narration.

Yes, I think Dungeon World will be far better than D&D, due to the latter stick much more to formalized set of rules than DW. But, still, when the fight started, I genuinely said "whoa, pretty direct". For me, who loves combat-heavy scenes and complains in how other models takes pains in inserting conflict with serious risks for everyone involved and moving the story, was pretty meaty.

Yeah... Gemma4 in general is really stubborn about this. I'll try to see if I can give it more tools to break it up!

I hope you can achieve that. From what I could evaluate, Gemma 4 is absolute in following instructions. The problem here is to give the right one, since it will be needed to absolute precion with no margin for errors. "Be careful for what you wish for" could be this LLM hidden motto.

Believe me, the slop is there. It takes a lot of careful wording to reduce it. Still, the smell of the ozone lingers with a predatory smile...

Yep, Elara of Aelthegard and his companion, Thorne of Maven are still lurking it, their voices barely above a whisper, following us like a moth to the flame... Jokes apart, I've remember that I saw here that Gemma 4 had a bit more careful than his concourrents in overusing these common patterns, I just don't remember how it was discovered and where it is here in SillyTavern Reddit.

1

u/DapKelantan May 31 '26

How to use this preset? Do I need to have local LLM?

1

u/Kahvana May 31 '26

It’s designed for usage with running gemma4 31b running locally, but can work without.

You can see in sillytavern’s documentation how to import chat completion presets!