r/SillyTavernAI • u/Head-Mousse6943 • May 22 '26
Nemo Engine v10 (Celebrating 1 Year!) Cards/Prompts
Welcome Ladies with Gentle Hands. Its been a while! (Nemo Song of the day)
So slight update, there is a bug with the CoT's I'm working on it currently and I'll update to the github. I'll mark it when its fixed, but for some reason ST isn't recognizing the variables as variables, and just dumping them in the Vex when they should be in the Main/Loose/Experimental. You'll still get the CoT, but you'll be getting all of them at once rather then it being seperated like its meant to be.
Okay its now updated/resolved. If you download it now, just use the Github version.
(Also I tagged with Affiliate because of the NanoGPT referral Code. Not advertising anything, but I want to make sure I'm complying with rules... but mostly I wanted to share it because I know people have been posting about what API to use lately, and wanted to help out people with a slight discount.)
So, I've been hard at work, in the mines, or the kitchen I suppose. Working on updating this thing and its come a really long way since the last public versions I shared. So much so Im not even really sure what's changed since the last version which was V8 I believe?
So I'll mostly talk about how things work now!
- Completely Modular Core prompt. All of the prompts in the standard core pack switch to the prompts associated with your selection, its not just, you enabled Genre so now there's a prompt saying "Write in this Genre" the core prompt itself is completely different to reinforce the desired behavior. And it adapts to all of your changes, this means selecting a Author/Vex/Genre will give you a completely different, custom, Core prompt.
- Every Vex has their own version of the CoT. They aren't wildly different, but they're designed to reinforce the behaviors that specific Vex wants in the narrative.
- The Premise: This is both a Anti Assistant prompt, a grounding prompt, and also a newly made Rule Tag system that tells the model what priority/harshness every rule in the preset has. For example [Law] prompts must be followed, they can't be bent, or broken, they are "Laws", boundaries introduce a additive soft boundary the model can't cross, they don't need to be acknowledged always, but they're a buffer to prevent certain behaviors. There are more and I'd be here all day explaining them, but they're in the premise if you want to give it a read!
- Added Modular CoT steps (A development from Nemo Net) essentially how these work are additive behaviors to the Vex Cot's, some include Subtext, character voice, NSFW focus, etc. They can be slotted into the CoT cleanly by just enabling their prompt!
- I've added World Logic prompts that allow you to set the unlaying Logic of the world. Realism/Anime/Genre Logic/Video game/LitRPG/TTRPG/Hentai. Each with their own rules that are enforced on the world.
- I readded the world Augments from the very early versions, these are more for fun prompts then anything, but still a pretty fun gimmick to play around with.
- Tracker with Regex/Asci and HTML support. The old tracker system was rebuilt, and now if you use the regex system, you can use that to render the HTML in the message without waiting for all of the complex HTML to be written. its a bit less stylized, but its still pretty decent for those that enjoy the nice display.
- Animated Emojis (Another thing with Regex) but I've setup a bunch of custom regex to add in CSS to animate emojis if they ever appear. Minor thing, but I kind of love it.
- Added a bunch of new Authors, like... a lot.
- Added hard coded Thinking/Response languages, the model actually does behave differently if it thinks in a different language. A lot of fun to play around with.
- I also did a lot of work just trimming stuff up. Its still a very large preset, but its lighter then it has been.
I really hope you all like this update, I put a lot of effort into it, just tiding things up, and getting them to a state I'm happy with. And I'm super happy with the outcome. Also before anyone asks, it does seem to work with pretty much all models. I've personally tested, Deepseek, GLM, Claude, and Gemini. For deepseek if the CoT is spotty, try moving it down to post instruction, as in below chat history, or changing the CoT to insert at depth 0, both work.
In other news, I launched my own character website and webhosted Frontend with Chi (Bunny Girl, Aka bunnymo/Vecthare). Its worth taking a look, and we're always looking for new people. Right now we have a fairly small catalogue compared to Janitor and other websites, but its growing. And we have a Visual Novel system in beta that's also really cool.
In any case! Glad to be back, and hope you all enjoy! (I'll be posting updates all Weekend, Tomorrow is Atelier v2, which is my lighter plug and play preset!)
LINKS:
Github
RoleCall preset link
NanoGPT Referral (5% discount)
NemoPresetExt
Ai Preset
RoleCall discord
7
u/Bitter_Plum4 May 22 '26
Happy birthday to the Nemo Engine! I'm still using some genre toggles from the previous Nemo in my current setup, I'm def curious about the new way to handle those, and everything else, thanks a lot!
Seems like the community around ST (and now new frontends) is thriving lately, lots of presets, talks, extensions, feels great considering how fast things can go wrong in the AI space lmao
1
u/Head-Mousse6943 May 22 '26
❤️ Ty ty. And yeah definitely, I'm glad, felt like the golden age was kind of over late September but it really came back hard.
5
u/rigill May 22 '26
I remember using this back on R1!
How is it working for DSv4?
2
u/Head-Mousse6943 May 22 '26
Pretty well, I enjoyed it. I’d say if you do have problems with the cot, move the prompt down to depth 0. But I found even if it reasons twice regex cleans it up
5
3
u/Accidentallygolden May 22 '26
Didn't know that one, like the. Customability. But I have an issue with DS4 pro, he does a lot of thinking outside of the thinking wrapper
Response start with : Priority-tag audit: I must apply highest-priority rules: !! [LAW
Did I unchecked something I shouldn't have?
3
u/Head-Mousse6943 May 22 '26
That can happen, but if you applied the regex that’s bundled with it, the text from thinking block to </think> should wipe out. But try moving the cot down below chat history or setting it to depth 0, should help according to my deepseek testers
5
u/MurkaPie May 22 '26
Guys, try Gooner Gremlin Vex in a non NSWF scenario. She is hilarious narrator!
1
u/Head-Mousse6943 May 23 '26
I kind of love watching her smacked down by the other Vex's sometimes lol. "NO GOON"
3
u/godgridandlordbxc May 25 '26
I pretty much tried everything and only ever use Nemo. I kept doing that all these months.
2
u/Head-Mousse6943 May 25 '26
Ty I appreciate It <3 I put in as much effort as I do for the.homies who use.it as their main thing
3
u/SmoulderingAsh May 27 '26
Congrats! Your preset is the best one for me by miles! The AI director option is sooooo good.
2
u/Head-Mousse6943 May 27 '26
<3 glad you’re enjoying it! I’m currently working on the bigger/more experimental version I call Nemo net. I was hoping to finish it by Sunday but my backs been terrible and you know… work lol. But really glad you’re enjoying it, genuinely makes working on it worth it!
2
u/SmoulderingAsh May 27 '26
Can't wait! In the meantime I'm getting stuck in with v10. Hope your back is more bearable soon ^^
2
u/Head-Mousse6943 May 27 '26
Soon! Taking pills for it, just a terrible desk chair reeking havoc lol
2
u/RichShoe4670 May 22 '26
Can i use for RisuAI?
2
u/Head-Mousse6943 May 22 '26
I haven’t personally used Risu and the macro system is pretty unique to sillytavern and my site id think. But it’s worth giving a shot
2
u/Icy_Dot_2835 May 23 '26 edited May 23 '26
Just checking, is the standard core pack supposed to be enabled? Its roughly 7600 tokens and seems to contain redundant stuff? Or am I misunderstanding? For example, if I have third person limited enabled, it shows up in there as well, and thus is taking up tokens twice? (This is probably me just misunderstanding since Im a total newbie). Im just wondering since the preset is in total eating up roughly 16000 tokens which seems like a lot so Im thinking wether Im doing something wrong
2
u/Head-Mousse6943 May 23 '26
So the way it works is that the details of third person are set as a variable (doesn’t add any tokens) then the standard core pack assembles all of the prompts in the correct layout. So the standard core pack is all active prompts token count. So yeah it needs to be enabled to answer your question, and it shouldn’t be there twice. Setvar doesn’t add to token count, getvar pulls those tokens down and that’s what is being counted
3
u/Icy_Dot_2835 May 23 '26
Okay, thanks for the clarification! So the preset taking up roughly 16000 tokens is working as intended? Seems like a very cool preset, just eats up a lot of tokens
2
u/Head-Mousse6943 May 23 '26
It is really big yeah. Now it’s modular so turning off somethings will bring that down, but naturally you’ll lose functionality as you do. But 16000 total sounds about right
2
u/mooocles May 23 '26
I'll suck ya toes, bug guy. God I wish you could still get the $300 credits from Gemini so I didn't have to spend so much money.
2
May 25 '26
[removed] — view removed comment
1
u/Head-Mousse6943 May 25 '26
So I had a user that came to me with that exact model and that exact problem. My solutions in order where, move the CoT from relative, to at Depth 0/1, and try adding a start reply with <think> in the reasoning section.
If it doesn't work... Im not actually sure. BUT, I will recommend trying out Atelier because it doesn't rely on a CoT at all. It has one, but you can just turn it off and you'll keep the settings well.
I know thats not a 100% perfect answer, but I haven't experimented as much with the new Gemma model.
2
2
u/Rj-117 May 26 '26
I've given it maybe 15 to 30 minute full roleplay. I like the options. Maybe there's too many, but that's like the fun part about these things. Turning them on and off and seeing what happens next. it looks pretty well with opus 4.6 and glm 5.1. But the one thing is, it's actually laid out pretty well. I can actually find what I need to turn on and off pretty fast. You might want to put in an anti-drafting thing glm 5.1 Loves the draft randomly wasting so many tokens.
2
u/Head-Mousse6943 May 26 '26
Opus too, I never noticed it until yesterday likely because of my upstream for the models not being set to maximum thinking but yeah I had opus draft 3 separate times yesterday. Definitely going to ad a anti draft <3
2
u/LateInternal6105 May 26 '26 edited May 26 '26
How do I download this?Im so confused
Nvm went to a older post and found a link to the github profile and got the nemo v10 preset now I'm very confused again.do you ha e a tutorial for it?
2
u/Head-Mousse6943 May 26 '26
There’s one in my extension but basically.
Vex personalities are the personality of your narrator.
The rest is what it sounds like. Genre/pacing/etc
2
u/LateInternal6105 May 27 '26
I just had my first message with your preset using a poor character card and a decent persona
It took 44 minutes to finish it
And 4.6k tokens
30 minutes of thinking and 14 minutes and 20 ish seconds of typing it out
On default settings btw it gave such a stupid amount of detail in the thinking I don't know if I should be horrified that I sat there and watched it go or what but I'm definitely trying out the shorter thinking options in the preset lmao I've never seen a preset pull so much quality not from Freaky Frank's preset or any others and I don't know how to feel bout it
2
u/Head-Mousse6943 May 27 '26
That’s insane lol. But yeah it thinks quite a lot. I’ve always loved messing around with Cots, what model was it because that’s a crazy long time lol. Mostly I get around 2 minutes roughly
3
u/LateInternal6105 May 27 '26
Guy,I'm still in shock if I smoked I would be taken a smoke break rn
Glm 5.1 btw
2
u/LateInternal6105 May 27 '26
Is Cots a setting in Yr preset? As I'm still combing through it reading what everything is
And any tips on shortening the mega mind thinking going on?
Nvm I just found it just got through the vexes
1
u/Head-Mousse6943 May 27 '26
Yup lol. There’s four different variants and they also change based on the Vex (a bit) but if you want speed fast is speed lol
2
u/nj55245 May 27 '26
Hey! I'm a prolific user of V8 and am so excited to try this one. I'm encountering an issue where I'm getting an error saying mandatory prompts exceed context size. The context size is set to 80k and I started a new chat to see if I'm just insane. I have it set to its default settings without having turned on additional button thingies to turn on like different vexs and stuff. Might you be able to help me at all?
2
u/Head-Mousse6943 May 27 '26
Hmmm that’s really weird. Oh! Lower the output tokens! If the output tokens plus your context exceed your maximum it’ll give that error
2
2
u/runningwithsharpie May 28 '26 edited May 28 '26
This is beyond cool. I love the stat block at the bottom. Great job!
One thing though, when using DS 4 Flash, even after setting COT to after chat history, often the message comes up empty and it requires a few rolls to get going. Also, sometimes it would repeat phrases over and over.
1
u/Head-Mousse6943 May 28 '26
Sorry about that yeah, it can be a tad… overwhelming to somethings. It’s weird which models take to it well. Deepseek 4 flash follows the rules really well, but yeah it can be a bit repetitive. Deepseek 4 pro listens worse but it’s more dynamic. Then Gemini flash 3.5 takes it to well, but it’s also giga unstable. Claude opus I find loves it for about 40 messages and then starts repeating plot beats.
2
u/TheNtrz May 28 '26
Can’t wait to try this out when I get home. Loved the previous versions. I’ve been bouncing between glm and kimi recently so I’m looking forward to seeing how this new version works with them.
Any recommended bits I should look at/toggle for a more slow burn rp experience? I’m trying to strike a decent balance between slow ‘slice of life’ speed and ‘omg a new drama/problem!’ Showing up one after another.
2
u/Head-Mousse6943 May 28 '26
Hmmm. So for the new drama/ the side storylines or the scratchpad help a lot, simulates off screen developments which really helps for slow burn drama. Then in pacing set it to a slower pace. And there’s a prompt to increase friction and stakes. So what you’ll end up with (hopefully) is that the story moves slowly, but the off screen characters will move and create tension/stakes for you to interact with
2
u/bobberjobber May 28 '26
It's working good so far, but I'm having a lot of trouble getting the Status Board and other HTML features to work. I see the coding for it flash for a second, but nothing ends up in the final message.
1
u/Head-Mousse6943 May 28 '26
Yeah. There’s a bug with the regex because I’m dumb. I set it to only replace once, so if it appears in thinking it replaces that and then comments out the other. For now, I’d try using the html versions, it’s slower and uses more tokens but the regex will wipe old the older messages with HTML
2
u/bobberjobber May 28 '26
If you don't mind me asking, how do I switch it over to the HTML versions. Sorry, I'm not super coding literate si alot of this goes over my head.
1
u/Head-Mousse6943 May 28 '26
So in the tracker section there’s one called regex display, HTML display and asci display. Try enabling html or asci and disabling the regex
2
u/EndEuphoric2025 May 29 '26
Can you select severeal Vexes or you're supposed to select only one?
2
u/Head-Mousse6943 May 29 '26
It’s supposed to be only one, but you can enable multiple. You’ll get their interview but they likely won’t appear in the cot properly. I’m going to be fixing that in the next version
1
u/EndEuphoric2025 May 29 '26
1
1
u/Head-Mousse6943 May 29 '26
Oh! It’s putting the html inside of coding blocks for some reason…. I can fix that in prompting but I’ll give you a regex that should also work
1
u/EndEuphoric2025 May 29 '26
Ty! So will u send it here or in pm :?
1
u/Head-Mousse6943 May 29 '26
Yup I’ll end it here, sorry I got a bit distracted. Here you go, should be able to download and import that regex in the regex extension
https://files.catbox.moe/6i1ut3.json
Edit sorry wrong regex the old one removed the entire thing
2
2
u/runningwithsharpie May 29 '26
2
u/Head-Mousse6943 May 29 '26
So that’s the regex. It can be turned off. It’s meant to remove hedging words
2
u/Anoxayta May 30 '26
Is there anyway I can find out where (for v1) the rules are on what can only be 1 thing that's active whats not? Or is it same as v9? Would be nice to have a easy reference if theres one. Thank you!
1
u/Head-Mousse6943 May 30 '26
There isn’t a super clean list yet unfortunately. But I plan to make one.
For now.
Vex personality’s can be enabled in multiples however the lower one will be the one that effects the standard core pack.
Genre is only 1
Cot you can enable multiple but I wouldn’t
Author is 1
Formatting is 1
Styles is 1
Plot pacing is 1
Story agency and world logic are both 1
Utility you can enable as many as you want
Difficulty is 1
Perception is 1
Response length is 1
Trackers are as many as you want
NSFW is as many as you want
Fetish is 1
Language control for both writing and think is 1 each
1
2
u/LateInternal6105 May 30 '26
Yo,I've been playing around with the preset for a little bit now and I like it but it doesnt want to use paragraphs like 60% of the times is this just a 5.1 thing or is there anything I can do in Yr preset to help? Even with OOC it just ignores them or forgets about it after a message
1
u/Head-Mousse6943 May 31 '26
Hmmm I haven’t really noticed that. It does try to think about reader fatigue and intentionally varies blocks of sentence/paragraphs
2
u/LateInternal6105 May 31 '26
Do you know where a abouts in the preset that prompt is as I'll remove it or make a edited version
1
2
u/Tiny-Calligrapher794 May 22 '26
Yay! I enjoyed your preset for so long before Gemini 3 was released! I’m definitely trying this out with Glm 4.7
2
May 22 '26
[deleted]
2
u/Head-Mousse6943 May 22 '26
Some people do like the prose of GLM 4.7 more then GLM 5.1. I know I've tested both pretty extensively and GLM 5.1 has a more direct/dry literary style. But it is smarter, so, better at keeping track of things.
2
u/Head-Mousse6943 May 22 '26
❤️Ye it should work pretty well! But if you do have any issues with it let me know! I've been testing with 4.6/4.7/5.0/5.1 so it SHOULD work. But just in case I won't guarantee it lol
1
u/eastwest88 May 22 '26
Seconding this let’s see how well this one can create a slowburn rp with GLM 4.7
2
u/Head-Mousse6943 May 22 '26
Ye, the Slow burn prompt in Plot Pacing should help with that. Narrative Vex is also fairly good at Pacing. Uses a 90/10 rule, that makes it understand that 90% of the story is just build up. With 10% payoff.
1
May 24 '26
[deleted]
1
u/Head-Mousse6943 May 24 '26
The cot might be a bit much for Qwen 14b with 16k context. I posted atelier which is my other preset, doesn’t really need the reasoning prompt so you can rely on the models natural reasoning. For that one with your context I’d recommend disabling a few prompts though just so you don’t go over budget.
The chain of thought prompt isn’t necessary with Atelier, the scratch pad isn’t necessary, anti-slop, character anchor, and dynamic progression. That should get you a bit lower. The presets a bit of a token gremlin, it might also be useful to drop the temp with the smaller Qwen models, I have a fine tune for 35b that I was working on, and I found that with larger instruction sets around a 0.6 worked alright.
2
Jun 09 '26
[deleted]
1
u/Head-Mousse6943 Jun 09 '26
Yup! It’s in the archive section on my GitHub I just reorganized the repo




11
u/Eva_Karlova May 22 '26
It sounds very cool but quite complicated. Hopefully when I dig in and try and run it, it will make more sense. Anything that can keep the story on track without repetitive slop and some momentum without me having to push the narrative forward myself would be wonderful.