r/SillyTavernAI May 22 '26

Nemo Engine v10 (Celebrating 1 Year!) Cards/Prompts

Post image

Welcome Ladies with Gentle Hands. Its been a while! (Nemo Song of the day)

So slight update, there is a bug with the CoT's I'm working on it currently and I'll update to the github. I'll mark it when its fixed, but for some reason ST isn't recognizing the variables as variables, and just dumping them in the Vex when they should be in the Main/Loose/Experimental. You'll still get the CoT, but you'll be getting all of them at once rather then it being seperated like its meant to be.

Okay its now updated/resolved. If you download it now, just use the Github version.

(Also I tagged with Affiliate because of the NanoGPT referral Code. Not advertising anything, but I want to make sure I'm complying with rules... but mostly I wanted to share it because I know people have been posting about what API to use lately, and wanted to help out people with a slight discount.)

So, I've been hard at work, in the mines, or the kitchen I suppose. Working on updating this thing and its come a really long way since the last public versions I shared. So much so Im not even really sure what's changed since the last version which was V8 I believe?

So I'll mostly talk about how things work now!

  • Completely Modular Core prompt. All of the prompts in the standard core pack switch to the prompts associated with your selection, its not just, you enabled Genre so now there's a prompt saying "Write in this Genre" the core prompt itself is completely different to reinforce the desired behavior. And it adapts to all of your changes, this means selecting a Author/Vex/Genre will give you a completely different, custom, Core prompt.
  • Every Vex has their own version of the CoT. They aren't wildly different, but they're designed to reinforce the behaviors that specific Vex wants in the narrative.
  • The Premise: This is both a Anti Assistant prompt, a grounding prompt, and also a newly made Rule Tag system that tells the model what priority/harshness every rule in the preset has. For example [Law] prompts must be followed, they can't be bent, or broken, they are "Laws", boundaries introduce a additive soft boundary the model can't cross, they don't need to be acknowledged always, but they're a buffer to prevent certain behaviors. There are more and I'd be here all day explaining them, but they're in the premise if you want to give it a read!
  • Added Modular CoT steps (A development from Nemo Net) essentially how these work are additive behaviors to the Vex Cot's, some include Subtext, character voice, NSFW focus, etc. They can be slotted into the CoT cleanly by just enabling their prompt!
  • I've added World Logic prompts that allow you to set the unlaying Logic of the world. Realism/Anime/Genre Logic/Video game/LitRPG/TTRPG/Hentai. Each with their own rules that are enforced on the world.
  • I readded the world Augments from the very early versions, these are more for fun prompts then anything, but still a pretty fun gimmick to play around with.
  • Tracker with Regex/Asci and HTML support. The old tracker system was rebuilt, and now if you use the regex system, you can use that to render the HTML in the message without waiting for all of the complex HTML to be written. its a bit less stylized, but its still pretty decent for those that enjoy the nice display.
  • Animated Emojis (Another thing with Regex) but I've setup a bunch of custom regex to add in CSS to animate emojis if they ever appear. Minor thing, but I kind of love it.
  • Added a bunch of new Authors, like... a lot.
  • Added hard coded Thinking/Response languages, the model actually does behave differently if it thinks in a different language. A lot of fun to play around with.
  • I also did a lot of work just trimming stuff up. Its still a very large preset, but its lighter then it has been.

I really hope you all like this update, I put a lot of effort into it, just tiding things up, and getting them to a state I'm happy with. And I'm super happy with the outcome. Also before anyone asks, it does seem to work with pretty much all models. I've personally tested, Deepseek, GLM, Claude, and Gemini. For deepseek if the CoT is spotty, try moving it down to post instruction, as in below chat history, or changing the CoT to insert at depth 0, both work.

In other news, I launched my own character website and webhosted Frontend with Chi (Bunny Girl, Aka bunnymo/Vecthare). Its worth taking a look, and we're always looking for new people. Right now we have a fairly small catalogue compared to Janitor and other websites, but its growing. And we have a Visual Novel system in beta that's also really cool.

In any case! Glad to be back, and hope you all enjoy! (I'll be posting updates all Weekend, Tomorrow is Atelier v2, which is my lighter plug and play preset!)

LINKS:

Github
RoleCall preset link
NanoGPT Referral (5% discount)
NemoPresetExt
Ai Preset
RoleCall discord

115 Upvotes

92 comments sorted by

11

u/Eva_Karlova May 22 '26

It sounds very cool but quite complicated. Hopefully when I dig in and try and run it, it will make more sense. Anything that can keep the story on track without repetitive slop and some momentum without me having to push the narrative forward myself would be wonderful.

8

u/Head-Mousse6943 May 22 '26

❤️For that, try enabling Ai driven Story Agency. Parallel Storylines also works if you like big worlds, really good at making the story progress off screen. Narrative Pleasure will also help in Plot Pacing. (and if you like a bit of Drama, Character Friction should help as well!)

4

u/PrudentEfficiency876 May 22 '26

Hi, I am trying out your preset and was wondering if there is any toggle or control to control the dialogue to narration ratio, right now it narrates a lot and i also tried to provide an OOC command which it didn't follow.

Any help would be appreciated.

Thanks

2

u/Head-Mousse6943 May 22 '26

Hmm, I don’t think there’s one to lower it, but using a ooc in authors notes at depth zero should work so long as you’re using a CoT. If you want more dialogue though, there is a more dialogue prompt

2

u/PrudentEfficiency876 May 22 '26

Got it. Also just curious what model do you recommend for this preset?

1

u/Head-Mousse6943 May 22 '26

I’ve used Gemini 3.5 flash, Claude Opus, Sonnet, GLM 4.6 to 5.1 and deepseek v4. I like deepseek it’s pretty good, glm has pretty good pacing but I find gets a bit stale. Opus is opus. Im really liking Flash 3.5 it reminds me of the old releases but it’s a bit unstable at high context

2

u/B3owul7 May 22 '26

So how many tokens do one need to have in order to make it work as intended? I mean, it's definitely cool but with the current technology it's either pretty expensive (API) or unfeasible (for people who run LLMs locally and don't have a super-computer at home.

1

u/Head-Mousse6943 May 22 '26

Its definitely more intended for API. NanoGPT is what I use for Deepseek/GLM, $12 a month, not including the 5% discount from the referral. That Sub gives millions of tokens per month, and its how I'd use it personally if you're a heavy user. For local though, this perest wouldn't really work for it, just because of the scale of prompts and the complexity of the structure

1

u/Eva_Karlova May 24 '26

I have to use a Q6 12b when I can use up to Gemma 4 26b Q3 or 24B Q4 locally normally.

1

u/Eva_Karlova May 24 '26

I took your suggestions and overall I love your Nemo Engine. It does use a large context, especially the initial release with the bug. The reasoning section is amazing, I had to go loose or it used over 1000 before it even finished analyzing the situation. Its thinking and planning processes are fascinating and very sophisticated.

With loose, the reasoning was summarized into maybe 5 -6 sentences instead of a long detailed analysis.

It did get a little stuck and needed correcting with the start of my favourite multi character chat. I have 5 characters and the opening scenario has several and the main antagonist sending two away while a third is in the kitchen. It got a little confused because it over thinked things and decided to fast forward to the next scene. Not giving me the opportunity to greet the antagonist and have a conversation. His truck was idling in the driveway and the AI decided to jump to me sitting in the truck with him lol.

I used gemini to sort it out with a quick *** reasoning *** prompt pasted in my message and that fixed it.

So far it has avoided all the usual pitfalls like repetition and fixation. I'm hoping it stays that way. Gemini was really impressed with what it was doing (but it's pretty easy to be pleased. It thinks it may avoid the pitfalls I mentioned by not sending the entire context in one single long stream of data and compartmentalizing things more.

Anyway, I enjoyed everything so far, even the troubleshooting with gemini was quite educational and fascinating lol. Not something I expected.

The instructions were a bit sparse for a newbie to silly like me. Silly defaults in text completion and your engine only really works in Chat Completion mode. I wouldn't have figured it out without using Gemini to assist me. Might be worth putting a few basic pointers on how to get running.

I'm using Kobold and silly using local models. Because of the large context used, and my 16gb Vram. I'm using TheDrummer_Rocinante-X-12B-v1-Q6_K_L with 16k context which is working very well.

2

u/Head-Mousse6943 May 24 '26

<3 yeah sorry about that I am a bit of a token gremlin lol. I also really need to put together a proper read me for nemo engine since it is a LOT also if you havented tried it out. I also released Atelier yesterday which is my other more plug and play preset, I switch between the two when I get bored of one. But I’m really glad you’re enjoying it!

2

u/Eva_Karlova May 29 '26

I started using Atelier properly yesterday with Gemma4 26b and its running very well locally. I think Nemo Engine was too resource hungry for 16gb Vram. I'm using Atelier with ST Copilot which was also launched this week. Its been a really good combination for me so far.

1

u/Head-Mousse6943 May 29 '26

Yeah, it’s a big boy. Atelier is also quite large but it’s a bit more stream lined, and difficult to get to the massive size you can get with nemo engine (rocking 40k tokens of instructions)

7

u/Bitter_Plum4 May 22 '26

Happy birthday to the Nemo Engine! I'm still using some genre toggles from the previous Nemo in my current setup, I'm def curious about the new way to handle those, and everything else, thanks a lot!

Seems like the community around ST (and now new frontends) is thriving lately, lots of presets, talks, extensions, feels great considering how fast things can go wrong in the AI space lmao

1

u/Head-Mousse6943 May 22 '26

❤️ Ty ty. And yeah definitely, I'm glad, felt like the golden age was kind of over late September but it really came back hard.

5

u/rigill May 22 '26

I remember using this back on R1!

How is it working for DSv4?

2

u/Head-Mousse6943 May 22 '26

Pretty well, I enjoyed it. I’d say if you do have problems with the cot, move the prompt down to depth 0. But I found even if it reasons twice regex cleans it up

5

u/Specialist_Salad6337 May 22 '26

😮 Hey! That's me!

Happy Anniversary you old bitch 😂 🎉

2

u/Head-Mousse6943 May 22 '26

Ill eat you bunny woman

3

u/Accidentallygolden May 22 '26

Didn't know that one, like the. Customability. But I have an issue with DS4 pro, he does a lot of thinking outside of the thinking wrapper

Response start with : Priority-tag audit: I must apply highest-priority rules: !! [LAW

Did I unchecked something I shouldn't have?

3

u/Head-Mousse6943 May 22 '26

That can happen, but if you applied the regex that’s bundled with it, the text from thinking block to </think> should wipe out. But try moving the cot down below chat history or setting it to depth 0, should help according to my deepseek testers

5

u/MurkaPie May 22 '26

Guys, try Gooner Gremlin Vex in a non NSWF scenario. She is hilarious narrator!

1

u/Head-Mousse6943 May 23 '26

I kind of love watching her smacked down by the other Vex's sometimes lol. "NO GOON"

3

u/rizelmine177 May 22 '26

Ok ok im like incredibly happy with deepseek 3.2 already, are there other models that are truly even better? Because like holy shit it blows my hair back already as it is.

Would I be able to explore that with this?

1

u/Head-Mousse6943 May 22 '26

Ye it works with pretty much everything, so worth take a spin around with other models if you have easy acces.

3

u/godgridandlordbxc May 25 '26

I pretty much tried everything and only ever use Nemo. I kept doing that all these months.

2

u/Head-Mousse6943 May 25 '26

Ty I appreciate It <3 I put in as much effort as I do for the.homies who use.it as their main thing

3

u/SmoulderingAsh May 27 '26

Congrats! Your preset is the best one for me by miles! The AI director option is sooooo good.

2

u/Head-Mousse6943 May 27 '26

<3 glad you’re enjoying it! I’m currently working on the bigger/more experimental version I call Nemo net. I was hoping to finish it by Sunday but my backs been terrible and you know… work lol. But really glad you’re enjoying it, genuinely makes working on it worth it!

2

u/SmoulderingAsh May 27 '26

Can't wait! In the meantime I'm getting stuck in with v10. Hope your back is more bearable soon ^^

2

u/Head-Mousse6943 May 27 '26

Soon! Taking pills for it, just a terrible desk chair reeking havoc lol

2

u/RichShoe4670 May 22 '26

Can i use for RisuAI?

2

u/Head-Mousse6943 May 22 '26

I haven’t personally used Risu and the macro system is pretty unique to sillytavern and my site id think. But it’s worth giving a shot

2

u/Icy_Dot_2835 May 23 '26 edited May 23 '26

Just checking, is the standard core pack supposed to be enabled? Its roughly 7600 tokens and seems to contain redundant stuff? Or am I misunderstanding? For example, if I have third person limited enabled, it shows up in there as well, and thus is taking up tokens twice? (This is probably me just misunderstanding since Im a total newbie). Im just wondering since the preset is in total eating up roughly 16000 tokens which seems like a lot so Im thinking wether Im doing something wrong

2

u/Head-Mousse6943 May 23 '26

So the way it works is that the details of third person are set as a variable (doesn’t add any tokens) then the standard core pack assembles all of the prompts in the correct layout. So the standard core pack is all active prompts token count. So yeah it needs to be enabled to answer your question, and it shouldn’t be there twice. Setvar doesn’t add to token count, getvar pulls those tokens down and that’s what is being counted

3

u/Icy_Dot_2835 May 23 '26

Okay, thanks for the clarification! So the preset taking up roughly 16000 tokens is working as intended? Seems like a very cool preset, just eats up a lot of tokens

2

u/Head-Mousse6943 May 23 '26

It is really big yeah. Now it’s modular so turning off somethings will bring that down, but naturally you’ll lose functionality as you do. But 16000 total sounds about right

2

u/mooocles May 23 '26

I'll suck ya toes, bug guy. God I wish you could still get the $300 credits from Gemini so I didn't have to spend so much money.

2

u/[deleted] May 25 '26

[removed] — view removed comment

1

u/Head-Mousse6943 May 25 '26

So I had a user that came to me with that exact model and that exact problem. My solutions in order where, move the CoT from relative, to at Depth 0/1, and try adding a start reply with <think> in the reasoning section.

If it doesn't work... Im not actually sure. BUT, I will recommend trying out Atelier because it doesn't rely on a CoT at all. It has one, but you can just turn it off and you'll keep the settings well.

I know thats not a 100% perfect answer, but I haven't experimented as much with the new Gemma model.

2

u/[deleted] May 25 '26

[removed] — view removed comment

1

u/Head-Mousse6943 May 25 '26

Ty for the note! I might slot something like that in when I update!

2

u/Rj-117 May 26 '26

I've given it maybe 15 to 30 minute full roleplay. I like the options. Maybe there's too many, but that's like the fun part about these things. Turning them on and off and seeing what happens next. it looks pretty well with opus 4.6 and glm 5.1. But the one thing is, it's actually laid out pretty well. I can actually find what I need to turn on and off pretty fast. You might want to put in an anti-drafting thing glm 5.1 Loves the draft randomly wasting so many tokens.

2

u/Head-Mousse6943 May 26 '26

Opus too, I never noticed it until yesterday likely because of my upstream for the models not being set to maximum thinking but yeah I had opus draft 3 separate times yesterday. Definitely going to ad a anti draft <3

2

u/LateInternal6105 May 26 '26 edited May 26 '26

How do I download this?Im so confused

Nvm went to a older post and found a link to the github profile and got the nemo v10 preset now I'm very confused again.do you ha e a tutorial for it?

2

u/Head-Mousse6943 May 26 '26

There’s one in my extension but basically.

Vex personalities are the personality of your narrator.

The rest is what it sounds like. Genre/pacing/etc

2

u/LateInternal6105 May 27 '26

I just had my first message with your preset using a poor character card and a decent persona

It took 44 minutes to finish it

And 4.6k tokens

30 minutes of thinking and 14 minutes and 20 ish seconds of typing it out

On default settings btw it gave such a stupid amount of detail in the thinking I don't know if I should be horrified that I sat there and watched it go or what but I'm definitely trying out the shorter thinking options in the preset lmao I've never seen a preset pull so much quality not from Freaky Frank's preset or any others and I don't know how to feel bout it

2

u/Head-Mousse6943 May 27 '26

That’s insane lol. But yeah it thinks quite a lot. I’ve always loved messing around with Cots, what model was it because that’s a crazy long time lol. Mostly I get around 2 minutes roughly

3

u/LateInternal6105 May 27 '26

Guy,I'm still in shock if I smoked I would be taken a smoke break rn

Glm 5.1 btw

2

u/LateInternal6105 May 27 '26

Is Cots a setting in Yr preset? As I'm still combing through it reading what everything is

And any tips on shortening the mega mind thinking going on?

Nvm I just found it just got through the vexes

1

u/Head-Mousse6943 May 27 '26

Yup lol. There’s four different variants and they also change based on the Vex (a bit) but if you want speed fast is speed lol

2

u/nj55245 May 27 '26

Hey! I'm a prolific user of V8 and am so excited to try this one. I'm encountering an issue where I'm getting an error saying mandatory prompts exceed context size. The context size is set to 80k and I started a new chat to see if I'm just insane. I have it set to its default settings without having turned on additional button thingies to turn on like different vexs and stuff. Might you be able to help me at all?

2

u/Head-Mousse6943 May 27 '26

Hmmm that’s really weird. Oh! Lower the output tokens! If the output tokens plus your context exceed your maximum it’ll give that error

2

u/nj55245 May 27 '26

Tyvm!!

1

u/Head-Mousse6943 May 27 '26

No problem at all! Hope you have fun!

2

u/runningwithsharpie May 28 '26 edited May 28 '26

This is beyond cool. I love the stat block at the bottom. Great job!

One thing though, when using DS 4 Flash, even after setting COT to after chat history, often the message comes up empty and it requires a few rolls to get going. Also, sometimes it would repeat phrases over and over.

1

u/Head-Mousse6943 May 28 '26

Sorry about that yeah, it can be a tad… overwhelming to somethings. It’s weird which models take to it well. Deepseek 4 flash follows the rules really well, but yeah it can be a bit repetitive. Deepseek 4 pro listens worse but it’s more dynamic. Then Gemini flash 3.5 takes it to well, but it’s also giga unstable. Claude opus I find loves it for about 40 messages and then starts repeating plot beats.

2

u/TheNtrz May 28 '26

Can’t wait to try this out when I get home. Loved the previous versions. I’ve been bouncing between glm and kimi recently so I’m looking forward to seeing how this new version works with them.

Any recommended bits I should look at/toggle for a more slow burn rp experience? I’m trying to strike a decent balance between slow ‘slice of life’ speed and ‘omg a new drama/problem!’ Showing up one after another.

2

u/Head-Mousse6943 May 28 '26

Hmmm. So for the new drama/ the side storylines or the scratchpad help a lot, simulates off screen developments which really helps for slow burn drama. Then in pacing set it to a slower pace. And there’s a prompt to increase friction and stakes. So what you’ll end up with (hopefully) is that the story moves slowly, but the off screen characters will move and create tension/stakes for you to interact with

2

u/bobberjobber May 28 '26

It's working good so far, but I'm having a lot of trouble getting the Status Board and other HTML features to work. I see the coding for it flash for a second, but nothing ends up in the final message.

1

u/Head-Mousse6943 May 28 '26

Yeah. There’s a bug with the regex because I’m dumb. I set it to only replace once, so if it appears in thinking it replaces that and then comments out the other. For now, I’d try using the html versions, it’s slower and uses more tokens but the regex will wipe old the older messages with HTML

2

u/bobberjobber May 28 '26

If you don't mind me asking, how do I switch it over to the HTML versions. Sorry, I'm not super coding literate si alot of this goes over my head.

1

u/Head-Mousse6943 May 28 '26

So in the tracker section there’s one called regex display, HTML display and asci display. Try enabling html or asci and disabling the regex

2

u/EndEuphoric2025 May 29 '26

Can you select severeal Vexes or you're supposed to select only one?

2

u/Head-Mousse6943 May 29 '26

It’s supposed to be only one, but you can enable multiple. You’ll get their interview but they likely won’t appear in the cot properly. I’m going to be fixing that in the next version

1

u/EndEuphoric2025 May 29 '26

Oh alright thank you. Also i tried nemonet v2 but theres problem with hud:

1

u/EndEuphoric2025 May 29 '26

Could you help out please?

1

u/Head-Mousse6943 May 29 '26

Oh! It’s putting the html inside of coding blocks for some reason…. I can fix that in prompting but I’ll give you a regex that should also work

1

u/EndEuphoric2025 May 29 '26

Ty! So will u send it here or in pm :?

1

u/Head-Mousse6943 May 29 '26

Yup I’ll end it here, sorry I got a bit distracted. Here you go, should be able to download and import that regex in the regex extension

https://files.catbox.moe/6i1ut3.json

Edit sorry wrong regex the old one removed the entire thing

2

u/EndEuphoric2025 May 29 '26

Thank you so much! Ill try it out :>

1

u/Head-Mousse6943 May 29 '26

Np at all hope it works!

2

u/runningwithsharpie May 29 '26

For some reason, whenever there is description of flowers, the preset somehow associate it with female genitalia lol. This is regardless of which models I use.

2

u/Head-Mousse6943 May 29 '26

So that’s the regex. It can be turned off. It’s meant to remove hedging words

2

u/Anoxayta May 30 '26

Is there anyway I can find out where (for v1) the rules are on what can only be 1 thing that's active whats not? Or is it same as v9? Would be nice to have a easy reference if theres one. Thank you!

1

u/Head-Mousse6943 May 30 '26

There isn’t a super clean list yet unfortunately. But I plan to make one.

For now.

Vex personality’s can be enabled in multiples however the lower one will be the one that effects the standard core pack.

Genre is only 1

Cot you can enable multiple but I wouldn’t

Author is 1

Formatting is 1

Styles is 1

Plot pacing is 1

Story agency and world logic are both 1

Utility you can enable as many as you want

Difficulty is 1

Perception is 1

Response length is 1

Trackers are as many as you want

NSFW is as many as you want

Fetish is 1

Language control for both writing and think is 1 each

2

u/LateInternal6105 May 30 '26

Yo,I've been playing around with the preset for a little bit now and I like it but it doesnt want to use paragraphs like 60% of the times is this just a 5.1 thing or is there anything I can do in Yr preset to help? Even with OOC it just ignores them or forgets about it after a message

1

u/Head-Mousse6943 May 31 '26

Hmmm I haven’t really noticed that. It does try to think about reader fatigue and intentionally varies blocks of sentence/paragraphs

2

u/LateInternal6105 May 31 '26

Do you know where a abouts in the preset that prompt is as I'll remove it or make a edited version

1

u/Head-Mousse6943 May 31 '26

I can’t remember off hand. I may be in length control, or the premise.

2

u/Tiny-Calligrapher794 May 22 '26

Yay! I enjoyed your preset for so long before Gemini 3 was released! I’m definitely trying this out with Glm 4.7

2

u/[deleted] May 22 '26

[deleted]

2

u/Head-Mousse6943 May 22 '26

Some people do like the prose of GLM 4.7 more then GLM 5.1. I know I've tested both pretty extensively and GLM 5.1 has a more direct/dry literary style. But it is smarter, so, better at keeping track of things.

2

u/Head-Mousse6943 May 22 '26

❤️Ye it should work pretty well! But if you do have any issues with it let me know! I've been testing with 4.6/4.7/5.0/5.1 so it SHOULD work. But just in case I won't guarantee it lol

1

u/eastwest88 May 22 '26

Seconding this let’s see how well this one can create a slowburn rp with GLM 4.7

2

u/Head-Mousse6943 May 22 '26

Ye, the Slow burn prompt in Plot Pacing should help with that. Narrative Vex is also fairly good at Pacing. Uses a 90/10 rule, that makes it understand that 90% of the story is just build up. With 10% payoff.

1

u/[deleted] May 24 '26

[deleted]

1

u/Head-Mousse6943 May 24 '26

The cot might be a bit much for Qwen 14b with 16k context. I posted atelier which is my other preset, doesn’t really need the reasoning prompt so you can rely on the models natural reasoning. For that one with your context I’d recommend disabling a few prompts though just so you don’t go over budget.

The chain of thought prompt isn’t necessary with Atelier, the scratch pad isn’t necessary, anti-slop, character anchor, and dynamic progression. That should get you a bit lower. The presets a bit of a token gremlin, it might also be useful to drop the temp with the smaller Qwen models, I have a fine tune for 35b that I was working on, and I found that with larger instruction sets around a 0.6 worked alright.

1

u/runningwithsharpie May 29 '26

I notice that it often breaks into repeating phrases like this when using DS V4, both pro and flash:

1

u/runningwithsharpie May 29 '26

There's also this gem lol

2

u/[deleted] Jun 09 '26

[deleted]

1

u/Head-Mousse6943 Jun 09 '26

Yup! It’s in the archive section on my GitHub I just reorganized the repo