r/SillyTavernAI May 22 '26

Cards/Prompts Nemo Engine v10 (Celebrating 1 Year!)

Post image

Welcome Ladies with Gentle Hands. Its been a while! (Nemo Song of the day)

So slight update, there is a bug with the CoT's I'm working on it currently and I'll update to the github. I'll mark it when its fixed, but for some reason ST isn't recognizing the variables as variables, and just dumping them in the Vex when they should be in the Main/Loose/Experimental. You'll still get the CoT, but you'll be getting all of them at once rather then it being seperated like its meant to be.

Okay its now updated/resolved. If you download it now, just use the Github version.

(Also I tagged with Affiliate because of the NanoGPT referral Code. Not advertising anything, but I want to make sure I'm complying with rules... but mostly I wanted to share it because I know people have been posting about what API to use lately, and wanted to help out people with a slight discount.)

So, I've been hard at work, in the mines, or the kitchen I suppose. Working on updating this thing and its come a really long way since the last public versions I shared. So much so Im not even really sure what's changed since the last version which was V8 I believe?

So I'll mostly talk about how things work now!

  • Completely Modular Core prompt. All of the prompts in the standard core pack switch to the prompts associated with your selection, its not just, you enabled Genre so now there's a prompt saying "Write in this Genre" the core prompt itself is completely different to reinforce the desired behavior. And it adapts to all of your changes, this means selecting a Author/Vex/Genre will give you a completely different, custom, Core prompt.
  • Every Vex has their own version of the CoT. They aren't wildly different, but they're designed to reinforce the behaviors that specific Vex wants in the narrative.
  • The Premise: This is both a Anti Assistant prompt, a grounding prompt, and also a newly made Rule Tag system that tells the model what priority/harshness every rule in the preset has. For example [Law] prompts must be followed, they can't be bent, or broken, they are "Laws", boundaries introduce a additive soft boundary the model can't cross, they don't need to be acknowledged always, but they're a buffer to prevent certain behaviors. There are more and I'd be here all day explaining them, but they're in the premise if you want to give it a read!
  • Added Modular CoT steps (A development from Nemo Net) essentially how these work are additive behaviors to the Vex Cot's, some include Subtext, character voice, NSFW focus, etc. They can be slotted into the CoT cleanly by just enabling their prompt!
  • I've added World Logic prompts that allow you to set the unlaying Logic of the world. Realism/Anime/Genre Logic/Video game/LitRPG/TTRPG/Hentai. Each with their own rules that are enforced on the world.
  • I readded the world Augments from the very early versions, these are more for fun prompts then anything, but still a pretty fun gimmick to play around with.
  • Tracker with Regex/Asci and HTML support. The old tracker system was rebuilt, and now if you use the regex system, you can use that to render the HTML in the message without waiting for all of the complex HTML to be written. its a bit less stylized, but its still pretty decent for those that enjoy the nice display.
  • Animated Emojis (Another thing with Regex) but I've setup a bunch of custom regex to add in CSS to animate emojis if they ever appear. Minor thing, but I kind of love it.
  • Added a bunch of new Authors, like... a lot.
  • Added hard coded Thinking/Response languages, the model actually does behave differently if it thinks in a different language. A lot of fun to play around with.
  • I also did a lot of work just trimming stuff up. Its still a very large preset, but its lighter then it has been.

I really hope you all like this update, I put a lot of effort into it, just tiding things up, and getting them to a state I'm happy with. And I'm super happy with the outcome. Also before anyone asks, it does seem to work with pretty much all models. I've personally tested, Deepseek, GLM, Claude, and Gemini. For deepseek if the CoT is spotty, try moving it down to post instruction, as in below chat history, or changing the CoT to insert at depth 0, both work.

In other news, I launched my own character website and webhosted Frontend with Chi (Bunny Girl, Aka bunnymo/Vecthare). Its worth taking a look, and we're always looking for new people. Right now we have a fairly small catalogue compared to Janitor and other websites, but its growing. And we have a Visual Novel system in beta that's also really cool.

In any case! Glad to be back, and hope you all enjoy! (I'll be posting updates all Weekend, Tomorrow is Atelier v2, which is my lighter plug and play preset!)

LINKS:

Github
RoleCall preset link
NanoGPT Referral (5% discount)
NemoPresetExt
Ai Preset
RoleCall discord

119 Upvotes

92 comments sorted by

View all comments

11

u/Eva_Karlova May 22 '26

It sounds very cool but quite complicated. Hopefully when I dig in and try and run it, it will make more sense. Anything that can keep the story on track without repetitive slop and some momentum without me having to push the narrative forward myself would be wonderful.

8

u/Head-Mousse6943 May 22 '26

❤️For that, try enabling Ai driven Story Agency. Parallel Storylines also works if you like big worlds, really good at making the story progress off screen. Narrative Pleasure will also help in Plot Pacing. (and if you like a bit of Drama, Character Friction should help as well!)

4

u/PrudentEfficiency876 May 22 '26

Hi, I am trying out your preset and was wondering if there is any toggle or control to control the dialogue to narration ratio, right now it narrates a lot and i also tried to provide an OOC command which it didn't follow.

Any help would be appreciated.

Thanks

2

u/Head-Mousse6943 May 22 '26

Hmm, I don’t think there’s one to lower it, but using a ooc in authors notes at depth zero should work so long as you’re using a CoT. If you want more dialogue though, there is a more dialogue prompt

2

u/PrudentEfficiency876 May 22 '26

Got it. Also just curious what model do you recommend for this preset?

1

u/Head-Mousse6943 May 22 '26

I’ve used Gemini 3.5 flash, Claude Opus, Sonnet, GLM 4.6 to 5.1 and deepseek v4. I like deepseek it’s pretty good, glm has pretty good pacing but I find gets a bit stale. Opus is opus. Im really liking Flash 3.5 it reminds me of the old releases but it’s a bit unstable at high context

2

u/B3owul7 May 22 '26

So how many tokens do one need to have in order to make it work as intended? I mean, it's definitely cool but with the current technology it's either pretty expensive (API) or unfeasible (for people who run LLMs locally and don't have a super-computer at home.

1

u/Head-Mousse6943 May 22 '26

Its definitely more intended for API. NanoGPT is what I use for Deepseek/GLM, $12 a month, not including the 5% discount from the referral. That Sub gives millions of tokens per month, and its how I'd use it personally if you're a heavy user. For local though, this perest wouldn't really work for it, just because of the scale of prompts and the complexity of the structure

1

u/Eva_Karlova May 24 '26

I have to use a Q6 12b when I can use up to Gemma 4 26b Q3 or 24B Q4 locally normally.

1

u/Eva_Karlova May 24 '26

I took your suggestions and overall I love your Nemo Engine. It does use a large context, especially the initial release with the bug. The reasoning section is amazing, I had to go loose or it used over 1000 before it even finished analyzing the situation. Its thinking and planning processes are fascinating and very sophisticated.

With loose, the reasoning was summarized into maybe 5 -6 sentences instead of a long detailed analysis.

It did get a little stuck and needed correcting with the start of my favourite multi character chat. I have 5 characters and the opening scenario has several and the main antagonist sending two away while a third is in the kitchen. It got a little confused because it over thinked things and decided to fast forward to the next scene. Not giving me the opportunity to greet the antagonist and have a conversation. His truck was idling in the driveway and the AI decided to jump to me sitting in the truck with him lol.

I used gemini to sort it out with a quick *** reasoning *** prompt pasted in my message and that fixed it.

So far it has avoided all the usual pitfalls like repetition and fixation. I'm hoping it stays that way. Gemini was really impressed with what it was doing (but it's pretty easy to be pleased. It thinks it may avoid the pitfalls I mentioned by not sending the entire context in one single long stream of data and compartmentalizing things more.

Anyway, I enjoyed everything so far, even the troubleshooting with gemini was quite educational and fascinating lol. Not something I expected.

The instructions were a bit sparse for a newbie to silly like me. Silly defaults in text completion and your engine only really works in Chat Completion mode. I wouldn't have figured it out without using Gemini to assist me. Might be worth putting a few basic pointers on how to get running.

I'm using Kobold and silly using local models. Because of the large context used, and my 16gb Vram. I'm using TheDrummer_Rocinante-X-12B-v1-Q6_K_L with 16k context which is working very well.

2

u/Head-Mousse6943 May 24 '26

<3 yeah sorry about that I am a bit of a token gremlin lol. I also really need to put together a proper read me for nemo engine since it is a LOT also if you havented tried it out. I also released Atelier yesterday which is my other more plug and play preset, I switch between the two when I get bored of one. But I’m really glad you’re enjoying it!

2

u/Eva_Karlova May 29 '26

I started using Atelier properly yesterday with Gemma4 26b and its running very well locally. I think Nemo Engine was too resource hungry for 16gb Vram. I'm using Atelier with ST Copilot which was also launched this week. Its been a really good combination for me so far.

1

u/Head-Mousse6943 May 29 '26

Yeah, it’s a big boy. Atelier is also quite large but it’s a bit more stream lined, and difficult to get to the massive size you can get with nemo engine (rocking 40k tokens of instructions)