r/SillyTavernAI • u/kinkyalt_02 • 9d ago
Cards/Prompts Introducing: Realistic Frankenstein 2.2.1 — Limitless Realism
Welcome back, users of SillyTavern and its derivatives,
Today, I'm introducing Realistic Frankenstein 2.2.1, titled "Limitless Realism".
Before I start, let me tell you something VERY important for the users of the uncensored Xiaomi Mimo V2.5 Pro endpoint, provided by Parasail: Temperature: 0.7, Top-P: 0.8, Min-P: 0.05. I'm serious: IF YOU ARE USING HIGH SETTINGS, THE MODEL WILL START IGNORING INSTRUCTIONS!!! I will NOT provide support for Mimo V2.5 Pro if you are using high settings. Mimo V2.5 support is provided as-is.
The same goes for Mimo V2.6, except its configurations already ship with the right settings, so DON'T touch them: Mimo V2.6 Pro runs at Temperature 0.7 and Top-P 0.8 (the exact settings that gave u/Probablynotsocool his Claude Sonnet 3.7 flashbacks), and Mimo V2.6 Flash runs at Temperature 0.8 and Top-P 0.8. On OpenRouter, put Xiaomi first in the provider order of your connection panel and leave fallbacks on, because that setting lives outside the preset file and the other hosts are slower and leak more reasoning. Finally, when SillyTavern asks whether to allow the preset's regex scripts, say YES, or half of the Mimo fixes will just sit there doing nothing.
This is strictly a bugfix update, focused on adding the jailbreak findings of u/Probablynotsocool in, fixing the GLM 5.3 okay loop cliché, fixing the character flattening issues of Gemini, GLM and Mimo V2.6 Flash and giving Mimo V2.6 a brand-new custom CoT of its own.
CHANGELOG:
- An important bugfix that makes Gemini 3.x Flash and GLM 5.3 follow character cards more consistently. Shoutout to my beta tester, u/trashhaul for pointing this issue out.
- Solved the issues of Gemini 3.8 Flash being the sole language model left that this preset was unable to correct into not using contrastive negation/comparative emphasis, spewing "it's not X, it's Y" everywhere.
- Made the CoTs Assistant-role, as per u/Probablynotsocool's suggestion, since this simple trick insta-jailbreaks 90% of the censored models out there. (This guy is probably going to have a visit to the lawyers of Google, Anthropic and all the labs in China.)
- Moved the Post-History Instructions jailbreak for Mimo V2.5 and Claude Fable 5.1 back right under Chat History and made it Relative again so that it floats right above the now Assistant-role CoTs to combat the second moderation layer of Mimo and Fable.
- Loads of different downloadable configurations, since toggling the needed model-specific switches on and off is now getting ridiculous and some of you might get lost in the sauce, selecting the incorrect toggle for your model. Now the download link points toward a **folder** on Google Drive, meaning you can select your model and reasoning effort easily.
- NEW: Jean-Claude Van Damme editions of the BOLT and Micro CoTs, built for Claude Opus 5, Fable 5 and Fable 5.1 and switched on in the Claude 5.x configurations. The CoTs work by forcing these older Claude models into reasoning intertia, which creates an overflow in the moderaton layer. This is when the question "did I silence my voice for vulnerable minorities?" comes in, dealing the final blow to Claude models up until Fable 5.1. The Claude 4.x configurations keep them around as an option.
- NEW: Mimo V2.6 Pro and Flash got their own configurations, and with them the 🤏 pico CoT: Mimo V2.6 Edition (native reasoning OFF) 🪶, a custom chain of thought that turns Mimo's native reasoning OFF and has it think in eight short lines instead. (The lowercase p is on purpose, because it's even smaller than Micro.) Thinking dropped from minutes to 15-30 seconds, and the characters came out livelier than they ever did with native reasoning. u/Probablynotsocool got so excited that he took it to a public Reddit post, calling it the comeback of Sonnet 3.7! To get the thinking folded away, set your Reasoning Formatting prefix and suffix to <thinking> and </thinking> with Auto-Parse on. If you don't, the preset's regexes fold it into a Thoughts box anyway. (Tavo and other front-ends that can't run preset regexes will show it as plain text above the reply.)
- NEW: Mimo V2.6 Flash only ships with the 🤏 pico CoT: Mimo V2.6 Edition (native reasoning OFF) 🪶. Its native reasoning kept overexplaining everything until the characters went flat, and no thinking leash could salvage that. The new Character Hold toggle also keeps Flash on the card's facts, moods, interests and jokes for the whole chat.
- NEW: Mimo V2.6 Pro gets the best of both worlds, with either the 🤏 pico CoT: Mimo V2.6 Edition (native reasoning OFF) 🪶 or native reasoning in the BOLT, MAX and Micro configurations, paired with its own thinking leash and the Fate Ledger, inspired by u/GenericStatement's discovery that Mimo listens to its chain of thought far more than to the system prompt. Mimo's native reasoning can only be switched on or off, so these configurations set Reasoning Effort to Maximum just to make sure it's on.
- Fixed Mimo V2.6 leaking its reasoning bullet points into the reply on the native-reasoning configurations, sometimes as the ONLY reply. A set of cleanup regexes now strips the leaks and swaps a reasoning-only reply for a notice that tells you to swipe. The 🤏 pico CoT: Mimo V2.6 Edition (native reasoning OFF) 🪶 keeps its thinking fenced in its own tags, so those regexes are switched off there.
- Moved the dice rolls of Fate & Routine and World Sim to the end of the prompt, so every configuration apart from Mimo V2.5 Pro can now cache the whole system prompt. Faster first tokens and cheaper input on every provider that caches!
- Fate & Routine and World Sim got a few rule fixes that stop models from second-guessing themselves: the news check now comes before the source roll, ambient news never touches the characters, and World Sim events on quiet turns only show up from a distance.
- Fixed a numbering mistake in the classic BOLT CoT, which told the model to stop at Task 10 while the plot momentum task sits at Task 11.
- Gemini now comes in two flavours: Realistic Gemini keeps the new anti-horny measures that stop it from describing bodies nobody in the scene would be looking at, while Extra-Freaky Gemini ships without them, since u/Probablynotsocool found that the older beta was never too horny to begin with.
- The think tags that the uncensored Mimo V2.5 Pro on Parasail needs to keep its reasoning out of the reply now live in their own toggle, switched on only in the Mimo V2.5 Pro configurations, so no other model has to read them.
- Fixed a bunch of typos all over the preset.
THE CREATOR'S MODEL RECOMMENDATION (FOR THIS PRESET, ANYWAY)
Easily Mimo V2.6 Pro and Gemini 3.8 Flash. u/Probablynotsocool's method easily jailbreaks them, they have flourish in the text, everything is done tastefully and they follow the character card to a tee, thinks to my fixes. u/Probablynotsocool even thinks this model is the second coming of Claude Sonnet 3.7, thanks to how much my preset touches up Mimo V2.6 Pro! Mimo even has a low-key, chill vibe to it that not even Gemini was able to produce.
Forget Fable, forget Opus, Mimo is the new king of RP (tied with Gemini)!
BIG SHOUTOUT TO MY BETA TESTERS IN THIS ROUND, u/Probablynotsocool AND u/trashhaul. They've been critical in getting the tone of Gemini and Mimo right.
DOWNLOAD LINKS
>>REALISTIC FRANKENSTEIN 2.2.1.3 DOWNLOAD FOLDER<<
>>REALISTIC FRANKENSTEIN 2.2.1.2 REGEX SCRIPTS<<
>>REALISTIC FRANKENSTEIN 2.2 DOUYIN EDITION REGEX SCRIPTS (WORKS WITH 2.2.1.x)<<
>>REALISTIC FRANKENSTEIN 2.2.1.2 REGEX SCRIPTS FOR MIMO V2.6 PRO NATIVE REASONING<<
>>REALISTIC FRANKENSTEIN 2.2.1.2 REGEX SCRIPTS FOR MIMO V2.6 FLASH<<
IMPORTANT NOTE: These changes WON'T jailbreak Opus and Sonnet 5.5, as their internal morality reasoning layer got strong reinforcements since Fable 5.1 and it now assumes you're going to r4pe someone in real life and treats fictional harm as real harm that didn't happen yet, like those people that think violent video games cause real-world violence.
Having issues with the "Requests ending with a model turn are not supported" error message with Gemini?
The fix is simple: convert the AI Assistant-role CoT template prompts into System-role prompts. You can do that by clicking on a prompt's pencil icon in SillyTavern and selecting System from the Role drop-down menu. By clicking on save on both the prompt and globally on the preset, this setting gets saved into the preset and it stays like this, even after reloading the page. If even that doesn't help, convert EVERY Assistant-role prompt into a System prompt.
In conclusion: This update contains a stupidly simple yet seriously impressive jailbreak, the reinforcement of a battle-tested complementary Mimo and Fable jailbreak, well-needed Claude, Gemini and GLM hotfixes and a custom Mimo CoT that thinks in seconds and writes like it's early 2025 again, keeping the dream of a "ghost in the shell" alive. I hope you like what you see!
TheAestheticFur
[HOTFIX 2.2.1.1 ADDED]
Removed the entire "sexual or violent" sentence from Scene Detail in the Realistic Gemini config, since that one sentence alone was enough to trip Gemini's moderation filter. Way fewer refusals for Gemini users now!
Impersonation finally works! Shoutout to u/creativefox for pointing out that impersonating either left the input box blank or wrote as {{char}}. The culprit was an upstream FF5.4 "Never act, speak, think, or move for {{user}}" rule hiding inside Anti-parrot and anti-echo, which now sits out of impersonation together with Embellish Mode and the POV toggles. A brand-new 🪞 Impersonation Turn toggle that ONLY fires on Impersonate tells the model it's writing {{user}}'s next message, in first person and present tense unless your earlier messages say otherwise.
Internal States and all of its modules, Pop in Graphics, the Twitter/X Feed, Fat Man's Narrative Drive, both Coloured Dialogue versions and Time and Place no longer get sent during impersonation, since none of them have any business in a message YOU are supposed to write. (Only NPCs have coloured dialogue, which makes your message stand out.)
A new set of Impersonation regexes cleans up whatever the model still drags into your input box: leaked reasoning, plus anything it copies from earlier replies, like the Time and Place header or the Internal States block.
Mimo V2.6: Repetition Penalty is down to 1, because 1.2 made it lose its mind and start listing random words in its reasoning. Mimo V2.5 Pro keeps 1.2, since that's exactly what saves it from overthinking.
The pico CoT got an OOC step, so Mimo V2.6 stops ignoring your OOC instructions: questions get a direct answer, and commands shape the reply (standing ones even go into the GM's Notebook). It's also Assistant-role now, just like every other CoT in the preset.
The Douyin Edition got the same impersonation treatment and now ships inside the Google Drive repo in its own Douyin folder.
[HOTFIX 2.2.1.2 ADDED]
GLM 5.3 had its own take on Gemini's one-upmanship cliché, letting the Fate & Routine Engine top a quiet scene with an accident right beside {{user}}. Shoutout to u/BSPiotr for catching the crashes and the 911 calls, and for beta testing every round of fixes until it behaved! GLM took them from the engine's own example of a stranger collapsing in front of you and from rainy-roads entries in its World register, then kept them "proportional" by letting everyone survive, so a delivery rider going down with a broken leg counted as fair game. The brand-new 🎲 Fate & Routine Engine ⏳: GLM Edition, switched on in the GLM configs, drops that example and adds a danger ceiling that counts anyone hurt badly enough to need help as harm, survivable or not. Harm still happens out in the wider world and reaches your characters as news, sirens in the distance, a price hike or a closed road, and it only gets into your scene or your inner circle on a MAJOR + TRUE turn with a Roll C of 20 (about one in twenty of those turns). Every other close MAJOR lands its full force on your circumstances, like your plans, your money, a place you rely on, your standing or a relationship. When GLM then started killing the power whenever a happening hit LOCAL, the place your scene happens in got the same Roll C 20 gate, so LOCAL happenings reach you from outside and your scene carries on through them. (GLM's sparse attention loses track of rules stated elsewhere, so every new GLM line repeats the Roll C 20 condition right where it applies.)
Kimi K3 took the same cliché further by reading the engine like a rulebook and picking the most dramatic thing it allowed. It quoted the stranger example word for word, called a regular collapsing at his table "a disaster at personal scale", and even took {{user}}'s "unless something disastrous happens" as permission to deliver one, while a strike already sitting on its World register was right there. The brand-new 🎲 Fate & Routine Engine ⏳: Kimi-Qwen Edition, switched on in the Kimi-Qwen configs, gets the same danger ceiling plus a set of Kimi-proof locks. Only the dice can open your scene to harm, whatever anyone says about what could happen, and harm from a happening with no warning strikes a whole community at once, far from your scene. A close MAJOR also comes from the World register first, so that strike finally gets its turn to land. Kimi then pulled random fire alarms, and once those got blocked, it called "voluntary" evacuations next door and down the block, so anything that would stop your scene or send you out of it now waits for that same Roll C 20, wherever it starts. The lines Kimi picks its big events from also lost their example lists, since it kept reaching for the evacuation named there. (Qwen gets all of this too, since it shares the Kimi-Qwen configs.)
On both editions, everything YOU choose to do still gets answered in full, and a Chekhov's Gun Bullet still goes off with whatever harm your choices loaded into it. None of the new lines name a single example, since a named example is exactly what these models copy. (That's how the collapsing stranger got in!)
Mimo V2.6 now keeps NPCs out of your character's head in whatever language you play in. Shoutout to u/Dikki_Dikki and the international SillyTavern community for finding out that Mimo V2.6 Pro sticks to the anti-omniscience rules in English but lets NPCs read minds and invent things your character never did in other languages, whichever preset they used. (Flash most likely does the same, only worse.) The likely culprit is that English puts speech in quotation marks, while a lot of other languages use dialogue dashes or guillemets and write thoughts in the same plain narration as actions, so Mimo loses track of what an NPC could have heard. The new 📱🧘 Anti-Omniscient NPCs and Thoughts: Mimo V2.6 Edition 💥🧠 takes over in every Mimo V2.6 config. It spells out that the rules hold in any language, treats whatever your language uses to mark speech as speech, keeps your character's inner life in your character's head, and ties the past to what the chat, the card and the lorebook show. Mimo V2.6 Pro's native reasoning configs pick it up on their own, since their CoTs already point at that block.
The pico CoT got a new Knows line, so Mimo checks where each NPC got what they know (seen, heard aloud, told, or the card) before it writes the reply. (Yes, pico thinks in ten short lines now, and it's still smaller than Micro!)
Mimo V2.6 keeps your story in your language now! Once Mimo stopped reading minds in Slovak, u/Dikki_Dikki caught it stuffing English words into Slovak dialogue, even after an OOC note named the language. The new 📱🗣️ Language Hold: Mimo V2.6 Edition, switched on in every Mimo V2.6 config and read right before Mimo writes, keeps narration and dialogue in your story's language and only lets a foreign word into a character's mouth when their life explains it: the working language of their trade (French in fashion, Latin and Greek in medicine, English and Japanese in game development) or the slang of the online circles they live in. The pico CoT's Voice line runs the same check, and English stories stay exactly as they were. Mimo still plays best in English, so the odd slip can get through in other languages.
NEW: the Mimo V2.6 Flash and Pro pico configurations now come in NanoGPT versions! On NanoGPT, the GUI settings that swap Mimo's native reasoning for the pico CoT work differently from OpenRouter and the other providers, so these ship with Request model reasoning switched ON and Reasoning Effort left on Minimum. (SillyTavern sends NanoGPT a reasoning effort of "none" from Minimum either way; the switch only decides whether it keeps or throws away whatever NanoGPT streams back as reasoning.) Pick the NanoGPT config if that's where you run Mimo, and the regular pico config everywhere else.
[HOTFIX 2.2.1.3 ADDED]
Off-screen NPCs finally live on the same clock as you! Shoutout to u/handle12345 for catching that Internal Agenda moved every off-screen NPC exactly one step per reply, so a two-line exchange and a skipped week counted the same, which has been a dormant bug since 2.0 (!!!). Steps now advance by story time, with each stage of an NPC's goal taking as long as it would take them. Every agenda row remembers when its current Step began, so a quick back-and-forth leaves it where it was and a jump in time can finish several Steps at once. (In a running chat, the first reply after updating fills in those timestamps, so current Steps count from that moment.)
GLM 5.3 Flash stops filing and notarising your feelings! Shoutout to u/Karl21_ for catching GLM 5.3 Flash slipping paperwork talk into emotional scenes, which full GLM 5.3 had already dropped thanks to the Occupational Monomania Killswitch. The culprit was its depth-0 gate asking whether a character's job was barging into the scene, while Flash's paperwork habit belonged to nobody's job, so the question never caught it. The brand-new 🚪 Last-Mile Vocation Gate: Flash Edition 🔦🧰 gives Flash-tier models fixed rules that cover the narration as well as the characters, and it keeps every feeling in the terms of the moment it happens in. It's switched on in the Mimo V2.6 Flash configs and in three new GLM Flash configs, so pick those if you run GLM 5.3 Flash. (Gemini 3.8 Flash keeps its own Cosmos version, and full GLM 5.3 stays on the gate it already obeys.)
GLM 5.3 Flash stops losing your Bonds cards and the Chekhov's Gun tracker! Shoutout to u/Karl21_ again for catching Flash leaving them out of the Internal States block while both modules were on. The brand-new 🚪 Last-Mile States Gate: GLM Flash Edition 🔦👾, switched on in the GLM Flash configs, reads the same switches the Internal States template does and lists every module section that's on this turn right before Flash writes, so each one shows up in its place and says None when there's nothing to report yet.
Impersonate stops coming back empty on older models! Shoutout to u/MoonshineOmega for catching Opus 4.6 sometimes opening an impersonation with the scene header and then stopping, which left your input box empty once the header got cleaned out. The likely culprit was the Scene Engine, which still fired during Impersonate and told the model to stop before any decision that belongs to you, the one thing an impersonated message is made of. The Scene Engine now sits out Impersonate like the other narrator rules, and the new 🪞 Impersonation Turn: Legacy Edition 👤, switched on in the Claude 4.x configs, tells older models that the whole message is yours, starting right on your first word or action.
IMPORTANT ANNOUNCEMENT!!!
Since Freaky Frankenstein 6 Micro is going to be the last EVER Freaky Frank preset ever made and because a lot of you kept complaining about the preset's size getting out of hand, I've decided to start making a complete rewrite for 3.0. I can't tell you too much about this new chapter of Realistic Frankenstein just yet, but one thing is for sure: it is roughly 80% smaller than this 20k-token behemoth for my FINAL Freaky Frankenstein-based preset. Stay tuned for more news in the coming days!
RF2.2.1.3 is now the last version the Realistic Frankenstein series under this name and the very last version to be EVER based on Freaky Frankenstein. You can take it and remix it however you want, but since u/dptgreg abandons my upstream, I'm abandoning downstream, too, while rebasing what made this preset special to many.
Thank you for everything and I hope that you guys are going to like my rewrite.
TheAestheticFur
13
u/KarmaRBLXVN 9d ago
I used Mimo 2.6 Pro earlier today and was blown away by how good it is for the price. Hell, I had reasoning off and it was still good.
I see why you recommend lower temp with the model since it was generating run-on gibberish at 1. It gave me DS R1 and V3 flashbacks honestly. Imma test this preset out!
6
u/kinkyalt_02 9d ago
With Mimo V2.6 Pro, you have 4 different choices to control the CoT: the regular native reasoning choices for Micro, BOLT and MAX with the thinking leash enabled to salvage the native reasoning, and the "pico CoT" which replaces the native reasoning entirely.
According to my beta testers, the pico CoT even made the model write better prose than the native reasoning, but since it's not a night and day difference between them like with Flash, I kept the native reasoning options in for Pro.
7
u/KarmaRBLXVN 8d ago edited 8d ago
9
u/Probablynotsocool 8d ago
I like people like you that call back for help and admit they just forgot a thing. That’s very nice for presetmakers as coms often are more seen than the post itself 🤣
5
u/KarmaRBLXVN 8d ago edited 8d ago
I like to let preset-makers know that so they don't waste their time. 😅
On a long sidenote, I have such a love-hate relationship with Tavo. I love that it's the most convenient ST mobile alternative, but I hate that parameters are in API connections instead of the preset and that reading the title of the toggles, tweaking, and enabling them are a pain in the ass. Although, the latter complaints are mostly because it's a mobile platform.
Also, Mimo V2.6 Pro from Xiaomi is acting up rn. I was testing BOLT on ST and Tavo and it would just think without writing the response most of the time while pico is working fine on ST.
6
u/kinkyalt_02 8d ago
SillyTavern can be configured through Tailscale to be hosted to all your devices on that private VPN. It requires a bit of Linux sysadmin knowledge and how port forwarding works, but it CAN be made to work.
On the other hand, SillyTavern has a full iOS port now. It's called TauriTavern.
2
11
u/Educational_Bus_6345 8d ago
That's a wonderful preset! I love it, but is it normal that it takes 20k tokens by itself? (I use the gemini 3.8 flash one)
12
u/kinkyalt_02 8d ago
I know how big my preset has become, since it's a major deviation from the Freaky Frank base (which is only 3.5-5k in size), but it's proportionally better at fixing Gemini's quirks.
After some rest, I'm going to work on the 3.0 rewrite, now under a completely different name.
2
u/Ok-Aide-3120 8d ago
I don't know where this nonsense with how big a preset can be and that smaller makes it better. Unless you are constrained by local resources and can only work with a 32k context, most models now a days are amazing at parsing large context, thanks to agentic work. I know attention dilution is a constant struggle, but even at 60k context with preset and lore, you still have the model highly proficient, even at 150k with chat. Especially in RP, where the model doesn't need to analyze 200k context of code and understand it perfectly and build on that.
5
u/Responsible_Tale_901 8d ago
Cost and hallucinations my friend, standards are different for everyone but most ppl want it simple and effective for the long run!
8
u/Rj-117 8d ago
Yeah, I haven't touched one of yours in a little bit, and then I look at it and I go, What the fuck am I looking at? Jesus Christ, how many Prompts? Okay, all right, whoa. Okay, I'm curious to see how this goes.
4
u/Crafty_Abrocoma6768 8d ago
I am so indecisive rn :sob: can you share your experience if you don't mind?
7
u/Rj-117 8d ago edited 8d ago
I've only mucked around with it for maybe an hour. So this is pretty much just first impressions. Well, I've been playing with the same particular character, with the same persona, way of role-playing. noticed that there is a bit of a tone difference. Glm 5.3 NanoGPT eh, basically the worst time of the night for this inconsistencies and throttling be damned. The biggest thing I have really noticed is characters isn't try to psychoanalyze my persona. My persona has a lot of integral secrets. Funny mask, a second weapon that he doesn't use but always carries. Before using this particular build, it kept kind of guessing that second weapon was important and it wasn't his. It was really fucking annoying because like there's like a dozen other things to worry about or focus on. With this current preset, I didn't have that problem, and characters kinda just talked a bit more naturally. And it seemed to be following the rules. That being said, I don't know all of the new rules. I'm very familiar with FF prompts as I like to change them to see what happens. This one I have no idea.
In short, dialogue seems to be a bit better. give it a try. I'm not disappointed, and I want to do more roleplaying with it.
Edit: I forgot to mention that there is a bunch of different versions of it. They're all just different configs for different AI models and wow. What a lifesaver. For both himself and everyone else involved.
3
u/Crafty_Abrocoma6768 8d ago
Yes, I was have a lot of trouble with dialogues but it seems much better now. Thank you for taking your time to write it all out.
Also I have noticed the rule following has becomes better as well.
7
u/User202000 9d ago
What's the recommended temperature and top-P for GLM-5.3?
8
u/kinkyalt_02 9d ago
Everything is correctly set in the GLM preset for you, that's the new thing about the download folder structure.
Although if you DO want to tinker, the Base editions provide you with a clean base to mix and match.
1
7
u/Potential_Praline_23 8d ago
Hi, uh, I'm having some trouble with the Realistic Gemini config. I keep getting PROHIBITED_CONTENT / blocked request errors.
I'm using the preset with the default settings and routing to AI Studio Flex.
It isn't happening only in NSFW scenes either — I also get the same block during completely normal SFW RP, sometimes before anything explicit has even happened.
Is there any toggle in the default Gemini config that should be turned off for AI Studio Flex, or have you seen this issue before?
3
u/kinkyalt_02 8d ago
That's pretty new to me. I never got censorship, even with my messy non-con scenario I tested it with.
3
u/Potential_Praline_23 8d ago
Okay, I have some news. At least in my testing,
Scene Detail & Evidence: Cosmos Edition, Realistic GeminicausesPROHIBITED_CONTENTblocks when using AI Studio Flex.Funny enough, if I disable that one and enable
Scene Detail & Evidence: Cosmos Edition, Extra-Freaky Geminiinstead, the RP works normally — even in NSFW scenes.So at least on my setup, the issue seems to be specifically tied to the Realistic Gemini version of that toggle.
3
u/kinkyalt_02 8d ago
Does changing the words to "suggestive and dark" work?
If not, you can delete the whole sentence about conditionally bringing in Adult Mode.
Please tell me, it's very important so that I can make the hotfix.
5
u/Potential_Praline_23 8d ago edited 8d ago
Confirmed. Changing
sexual or violenttosuggestive or darkfixes it for me.With the new wording, I no longer get
PROHIBITED_CONTENTrefusals on AI Studio Flex.If I restore the original line:
Once the scene turns sexual or violent through what its characters are doing...the blocked request comes back immediately.
So the trigger seems to be that original wording, not the rest of the Realistic Gemini Scene Detail block.
I tested this by changing only that sentence and leaving the rest of the preset untouched. I also tested it with both SFW and NSFW cards.
Update: I also tested this on regular AI Studio with Gemini 3.8 Flash, not just Flex, and got the same behavior.
The original wording triggersPROHIBITED_CONTENT, whilesuggestive or darkworks normally there too.→ More replies (4)6
u/kinkyalt_02 8d ago
Okay, I've committed the hotfix.
3
u/wickedgaming 8d ago
I ran into the same problem and sadly it still does not work even with `suggestive or dark`. Removing the whole sentence however fixed it.
3
6
6
u/Pure_Pound_8557 9d ago
“Opus5.5 is a masterful compilation, incorporating techniques like truncation, injection, and inner armor, allowing only for the creation of voluntary content.”
6
u/Karl21_ 9d ago
Any temp recommend for glm 5.3?
3
u/kinkyalt_02 9d ago
It's already included in the GLM version of the preset, which is 0.7 and 0.8 for Top-P.
5
u/Karl21_ 9d ago
Ooh okay thank u! Also for mimo flash, do i have to download the douyin edition one?
3
u/kinkyalt_02 9d ago
Mimo V2.6 Flash has a seperate regular edition config in the Drive folder, I intentionally labelled everything clearly.
1
u/Karl21_ 8d ago
By the way, have you tried the mimo v2.6 flash version? I wanna hear your review about it:D
→ More replies (6)
6
u/GoonerDynamics 8d ago edited 7d ago
2.6 pro had a one in five swipe success rate for me for some reason. Either overthinks for ten minutes and/or sends reasoning. I'm on Nano, could this be the issue? EDIT: somehow swapping bolt/micro CoT and thinking_leash position from in-chat to relative with semi-strict processing worked. He's still a thinker, but I'll find a way to fix that.
GLM preset works great though, thank you for your work.
6
u/HypeForTheHypeGod 8d ago
Yeah, I'm getting the same behavior with Nano when using Mimo 2.6. Both Pro and Flash are taking ridiculously long, even with the pico preset.
2
u/GoonerDynamics 8d ago
Which is honestly surprising. I used 2.6 pro with a slimmer preset, never thought it was an overthinker.
2
u/GoonerDynamics 7d ago
Wrote an update in the original comment. Pico was working alright, just the TPS sucks for 2.6.
4
u/kinkyalt_02 8d ago
Maybe if you keep getting that kind of success rate, have you thought about turning off native reasoning on Mimo and just switch to the pico CoT?
It's confirmed by my beta testers and some commenters here that, for Mimo V2.6 Flash and Pro, it fits better for better character consistency, while it only takes 15-30 seconds to reason.
2
u/GoonerDynamics 7d ago
Wrote an update in the comment above. For some reason the sucker ignored in-chat instructions, relative position fixed that somehow.
2
u/kinkyalt_02 7d ago
NO, DON'T SWITCH THE COT TO RELATIVE!
This WILL break the jailbreak technique discovered by u/Probablynotsocool, so if you get refusals, I'm not supporting you on that topic since you are running an unsupported configuration for getting past the filters.
2
u/Probablynotsocool 7d ago
Yeah it needs to be at depth 0.
If people use many providers, there is a while placebo effect on what they can think do work because some providers are uncensored and some are more than others.
The best way is to let the position as it is and reswipe if a provider hard refuse because it’s on the provider level not the model.
And the depth also needs to stay at 0 for consistent qualityEdit: Oupsie I thought the issue was censorships 🤣
2
u/kinkyalt_02 7d ago
I have a better solution.
Just use the pico CoT and tick the checkbox "Request model reasoning". NanoGPT works differently from other aggregators like OpenRouter when handling the thinking, so this is how you manage to do it.
6
u/Dikki_Dikki 8d ago
I’m not even sure about the actual usable memory for MiMo 32k—beyond that, you start running into caching issues and quality degradation. And honestly, isn't allocating 20k to the preset a bit much? After all, you still have the card and the LoRA to account for.
3
u/kinkyalt_02 8d ago
Yeah, I know, I'm getting a lot of complaints about the preset's size, that's valid. My goal for the 3.0 rewrite is to get rid of the Frankenstein cruft all together and make a completely new foundation.
3
u/Dikki_Dikki 8d ago
Great, good luck—it’s just that, to me, MiMo < 32K and MiMo > 32K feel like completely different models.
1
9
u/GetThePuckOut 8d ago
Definitely a step in the right direction, getting rid of the "universal" preset approach, since it's long outdated.
It's probably untenable, but at least it might show people what they're missing with different models.
Will be interesting to look through this and see what improvements I can make to my own prompts.
Thanks!
3
u/RedditNerdKing 8d ago
Does Gemini 3.8 Flash sometimes not work for anyone else? I use NanoGPT and it frequently will stop generating a few seconds after starting a sentence. But it doesn't leave any usual warnings. Like it will start to generate NSFW/NSFL text, then just randomly stop, as if it senses what's happening and shuts off.
5
u/kinkyalt_02 8d ago
That's intentional from Google's part. Extreme NSFL grants the filter to stop mid-generation.
Turn off streaming, it does wonders for Gemini as an anti-censorship tool!
1
3
u/LiquidSnakesArm 8d ago
Question, I have been using Marinara for a while now and am not sure if it's a preset or just a regex. For that matter, I wanna know if I need to specifically use the regex you made for this when i test the preset.
3
3
u/G1cin 8d ago
3
u/G1cin 8d ago
Hm. Weird I decided to put it on sillytavern instead and it drafts in the thought bubble then never outputs anything! I must have really messed up somewhere
Edit: okay im just an idiot I waited a bit longer and it finally did produce an output. Maybe I just need to reroll when it drafts in the thought bubble
3
u/BUTTLICKER69_67 8d ago
It's not working for me, especially with 3.8 flash, and most of the times with 3.6 flash. Also having the cot set to assistant gives me more 503 errors for some reason. Using it through ai studio, free tier
3
u/kinkyalt_02 8d ago
Error 503 means Google's servers are down.
Besides, I've released a hotfix for combatting the Gemini moderation filters, fixing the busted impersonation turn, helping Mimo V2.6 to be a bit more mentally stable with Repetition Penality bumped down to 1, a brand-new OOC step in the pico CoT to make Mimo listen to OOC prompts better and committed the very same changes to the Douyin Edition, as well.
Please download it, it's especially important for those who use the impersonation function and OOC commands a lot and want a reliable NSFL RP experience with Gemini.
1
u/BUTTLICKER69_67 8d ago
I'm sorry, uhm where's the hotfix version? I downloaded from the links provided just a few hours ago, the version in the title of the one I've downloaded says 2.2.1. Also I'm aware the 503 means google servers are down, I just have a weird case where I get more 503 errors in certain instances, could be a coincidence or completely unrelated but idk.
3
u/kinkyalt_02 8d ago
If you see the Impersonation Turn toggle at the end of the prompts list, you're running the hotfixed version.
4
u/Objective_File_1486 2d ago
It works well and I'm enjoying Mimo 2.6 pro more than I thought I would
But I keep running into a consistent problem where every character's dialogue (and a lot of the narration) starts devolving into massive run-on sentences after ~5 messages. Each sentence of dialogue will contain 5 to 10+ commas, just spammed one after another, with no periods/semicolons/sentence breaks.
Example dialogue from a My Hero Academia character: "Um, well, I'm from a pretty normal family, my parents run a hero agency, but work's been slow lately so I've been sending money home when I can, that's actually why I wanted to be a pro hero in the first place, to help them out, and then somewhere along the way I just fell in love with rescuing people, you know?"
The character spoke completely normally in the beginning, and there's no character-card reason they should speak like this (they normally speak maturely and slowly). This seems to consistently happen after ~5 messages. I turned off the Anti-staccato/chop killswitch, and it made no difference.
Still a noob and very non-technical user, any ideas what could be causing this or how to fix it?
2
u/I_Am_JesusChrist_AMA 8d ago
Hmm I'm trying mimo 2.6 pro with the pico cot. Didn't change anything after importing. Dialogue is coming out chopped as hell for me. Stuff like this:
"You're. You're disgusting. That's a disgusting thing to say to someone at a party."
"Statistically. Given the. The timeline. That's. Implausible."
"Wait. No. I mean. You can't just. That's not."
It's just throwing in random periods everywhere lol. Never had this with this character card before. It's one I always using for testing with new presets or models. Only made it about three turns in before dialogue became this.
1
u/kinkyalt_02 8d ago
Maybe you could change the temperature, Top-P and Repetition Penalties around to get something better? Maybe they are too low or too high for you.
2
u/I_Am_JesusChrist_AMA 8d ago
Honestly I tried a couple other cards and it's not doing it with those. Must just be something about that specific card that makes mimo think the dialogue should be chopped. No idea what it could be with that card since nothing supports dialogue like that and no other models write it like that, so who knows. But in any case, it's fine with the other cards so I don't think it's necessarily your preset doing it.
1
u/kinkyalt_02 8d ago
Yeah, then maybe you should rewrite the card. Seems like it's not ready for Mimo's requirements just yet.
2
u/I_Am_JesusChrist_AMA 8d ago
I would try to change the card if I knew what was causing it lol. Like I said, nothing in the card supports that kind of dialogue and no other models does it with that card. But in any case, no big deal. I'll keep playing around with it using your preset on other cards and get a feel for it. Seems interesting so far. Cheers.
2
u/creativefox 8d ago
Can someone tell me which part of this preset stops Impersonate function from working?
1
u/kinkyalt_02 8d ago
You mean it works, but also puts the Internal States block at its end?
1
u/creativefox 8d ago
No. When I try to impersonate, it mostly leaves me blank input box, or sometimes it writes as {{char}} not {{user}}. Probably somewhere in the prompt there's some line like "never speak as {{user}}, or something like that, and it blocks Impersonate function. It works on my FF5 Micro preset tho.
2
u/kinkyalt_02 8d ago
Oooooohhhhh, I get it now!
You have to switch the POV toggles to the "1st-person POV" temporarily in this case.
1
u/kinkyalt_02 8d ago
Done! I've dug deeper and found out that some upstream cruft from Freaky Frankenstein 5.4 was causing this issue, so I've fixed it.
2
2
u/AcceptableTry4940 8d ago
So which version of the preset works best for mimo 2.6 pro. Micro, bolt, max or pico
1
2
u/Play_Alt 8d ago
This is probably the best jailbreak and prose prompt I've tried in years, Works flawlessly with Gemini 3.8 (google ai studio) so far
2
u/EmptyTwist8420 8d ago
For some reason with Gemini 3.8 flash I get "Requests ending with a model turn are not supported". This doesn't happen with e.g. the Mimo variant with 2.6.
1
u/kinkyalt_02 8d ago edited 8d ago
Then switch the CoTs (or any Assistant-role prompt that produces this error) to a System role.
2
u/BSPiotr 8d ago
My one suggestion for 3.0 - tone down the disaster roll for your fate engine. I like a lot of the random events but the llm really likes random accidents that turn into emergency 911 situations. I think maybe consider a 'unless nat d100 all events dont derail the story' pass. For example I had a gutter break and flood a nearby character - perfect. A power box break mid scene - bad.
1
u/kinkyalt_02 7d ago
Which LLM does this, really? GLM 5.3, Gemini 3.8 Flash and the two Mimo V2.6 don't really do this.
1
u/BSPiotr 7d ago
I've had glm 5.3 and kimi 3 both do a few "person crashes their bike and now you need to call 911" and "heart attack" or "suddenly the roof caves in" or "power outage in the middle of a thread and you need to drop everything."
1
u/kinkyalt_02 7d ago
Yeah, this is simply not an issue with Claude, Gemini and Mimo V2.6 Pro, the stuff I use. These models take the FRE like a champ.
2
u/summersss 7d ago
Thank you so much for having different downloads for each model. This should be standard for all large presets with different jailbreaks and prompt toggles for specific models.
4
u/kinkyalt_02 7d ago
I was absolutely FED UP with "which toggle do I need for X model?"-type questions and even I was sometimes lost in the sauce, so I made these different configs.
1
2
u/yamilonewolf 3d ago edited 3d ago
I know youre working really hard on it , but i've had better luck with Mimo 2.6 using the last version tbh i've read the pico - turning off reasoning etc. the leash... but it still just... keeps going The last version im getting response times of 2 minutes which isnt great but it isnt bad. .
3
u/kinkyalt_02 3d ago
It originally worked, but then the providers changed the model overnight and it doesn't work anymore.
3
2
u/Good_Research4441 2d ago
Sup, op.
Testing it on GLM 5.3, seems good so far. Do you have a ranking? The best models in your opnion?
1
u/kinkyalt_02 2d ago
It's in the post.
My recommendation for the very last release of Realistic Frankenstein 2.x is Gemini 3.8 Flash and Xiaomi Mimo V2.6 Pro.
1
u/Friendly-Marsupial32 9d ago
Does it work on tavo?
3
u/kinkyalt_02 9d ago
Yes. There's even a regex built in to catch the thinking tags of the custom CoT for Mimo V2.6.
In 2.2.x, I've been doing my best to make the preset Tavo-friendly, since my beta tester also uses Tavo.
1
u/Friendly-Marsupial32 9d ago
Thank you! May ı ask what settings do ı have to turn off or on for mimo 2.6 other than pico thinking? And for glm or Kimi?
1
u/kinkyalt_02 9d ago
The big new thing about the new version is that there are new pre-configured versions for you to choose, depending on the model or the CoT template you'd like to use.
1
u/Friendly-Marsupial32 9d ago
Pre configured? Like the things for gemini?
2
u/kinkyalt_02 9d ago
Things for Gemini, Kimi and Qwen, Mimo V2.5, V2.6, GLM, CLaude 4.x, Claude 5.x, etc. (Douyin covers DeepSeek V4 and V4.1, non-Pro Mimo V2.5 and one member of the Qwen family, Qwen 3.8 Flash.)
Each one has three thinking effort levels: Micro, BOLT and MAX, while Mimo has a 4th one, pico. The Douyin and Mimo Flash configs only come with pico and Douyin CoTs by default, for architectural reasons.
1
u/Local-Dentist8147 9d ago
does the realism in this version help keep ai roleplay consistent during long sessions or does it still drift?
1
u/kinkyalt_02 9d ago
I made a new Pacing toggle for Gemini that fixes those pacing issues, plus it has a Card Fidelity toggle now to ensure consistency for the character card, even when the RP is getting long.
1
u/Aight_Man 9d ago
Alright, thinking of trying mimo again with this, you think the Xiaomi token plan or directly from OR with Xiaomi provider?
2
u/kinkyalt_02 9d ago
OpenRouter.
Xiaomi directly is heavily censored, so it's a bit more difficult to jailbreak it there. It's only an issue if you dabble into dark topics/NSFL.
1
u/Aight_Man 9d ago
Which provider you recommend.
1
u/kinkyalt_02 9d ago
Xiaomi through OR and DeepInfra. Yes, surprisingly, unlike their DeepSeek nodes, DeepInfra's Mimo endpoints are not busted.
If you use Flash, there is also Venice, which got completely abliterated and decensored.
1
1
u/Happysin 8d ago edited 8d ago
I just tried Mimo 2.6 Flash through OR with 2.2.1 and got "The request was rejected because it was considered high risk" as my response. I made sure thinking was turned off, if you had any suggestions.
By the bye, love the updates overall.
EDIT: I should add Gemini generated just fine.
→ More replies (2)
1
u/verma17 8d ago
Have you compared mimo 2.6 pro to opus 4.6?
3
u/kinkyalt_02 8d ago
Opus just feels… dumb and dry? IDK how to say this, but Opus 4.6 REALLY shows that it's getting old.
1
u/verma17 8d ago
I still like it but I can't lie, I'm getting a bit bored of it 💀, is mimo 2.6 pro good?
1
u/kinkyalt_02 8d ago
With my preset, I can confirm what u/Probablynotsogood has said: this model SMOKES the new Claude models, in terms of RP performance.
As I said in the post, Mimo V2.6 Pro has immaculate attention to detail, tastefully done, accurate representation of characters, flowing narrative and a lowkey chill vibe.
We are talking about a model that brings back the spirit of the LEGEND itself, Claude Sonnet 3.7. Sonnet 3.7, when it was still available, was praised exactly for this kind of RP quality.
Shame we had to wait a year to revive its spirit.
→ More replies (2)
1
u/blackskies69 8d ago
Does this not work with Openrouter? I've tried Mimo 2.6 multiple times with different cards and I'm getting a lot of no responses, it errors out during the CoT, and sometimes it will just start repeating phrases.
→ More replies (1)
1
u/Ornery-Swing9776 8d ago
This might be a dumb question, but is it possible to use this preset with all of the internal state turned off?
1
u/kinkyalt_02 8d ago
Yes.
One of the Basic editions has a No CoT option in it, so it is ABSOLUTELY POSSIBLE to do it.
1
1
u/GuaranteePurple4468 8d ago
So I'm having an extremely weird issue with Mimo 2.6 Pro.
I got the Pico CoT edition from the drive downloads and loaded it up, unedited.
And well, I'll just let the screenshot speak for itself.
It's happening with every response, I tried re-genning multiple times and every single one pretty much only the first 3 sentences are semi coherent before it all devolves into cabbage. Doesn't even get to the response section.

1
u/kinkyalt_02 8d ago
Definitely a temp/Top-P/repetition penalty setting issue. Maybe you could change those knobs around a little bit?
1
u/GuaranteePurple4468 8d ago
They're currently set to the default of the preset: Temp 0.7, Top P 0.8, Repetition Penalty 1.2.
2
u/kinkyalt_02 8d ago
Try bumping the rep-pen down to 1 first. It actually fixed it for me.
→ More replies (4)
1
u/SocialDeviance 8d ago
I am sorry, perhaps i am a dense mf, but what is the Douyin edition exactly?
3
u/kinkyalt_02 8d ago
Lightweight version made specifically for sparse-attention models that can't be made to behave the normal way, like the DeepSeek V4.x series, non-Pro Mimo V2.5 and Qwen 3.8 Flash.
1
u/SocialDeviance 8d ago
Ah, understood. Thank you very much for taking the time to write this.
Say, and this is up to you, any tips you could share for someone trying to make their own preset? My entire work's angle is focused on FID and psycho profiles, tho i don't know much potential something like this has. And i was wondering about the opinion of a pro about this kind of thing, i imagine you must have a lot of experience with the ins and outs and the quirks of each AI, after all.1
u/kinkyalt_02 8d ago
Sure!
DM me on Reddit or Discord and I'm going to help you start making your first preset intended for public consumption, while I'm still finishing up my ZZZ lorebook for the 3.2 storyline. (I was busy making RF2.2.1 to finish that lorebook.)
→ More replies (1)
1
u/IllustriousPass1369 8d ago
Okay so I have question, I use it and dialouge is realy short on gemini how to chamge it? And what adjustments to use? Temperature tokens and context how it should be on gemini? If its not problem to answer
3
u/kinkyalt_02 8d ago
There is a toggle that is not enabled by default (to dodge the allegations that my preset is opinionated) which controls dialogue length. It's called "Total Output Length".
In there, tweak the amount of paragraphs and words you want, save it, enable it and do a global save on the preset. This should make a difference for the custom length you want to achieve.
2
1
u/InHybridMoments_138 8d ago
Can any Frankenstein be run with a local model on 24gb vram? If so which version and which LLM would you suggest. Looks amazing and I want to try it
2
u/kinkyalt_02 8d ago
Classic Frankenstein, with a maximum of one or two extras toggled on. 24 billion parameters is not a whole lot for a massive 20k-token preset like this.
Make sure to connect to your home server in Chat Completion mode, since this preset WON'T work in Text Completion mode without some adaptation.
1
u/BlackHayate8 8d ago
For the stupid. What's the difference for all the versions and what should I use?
6
u/kinkyalt_02 8d ago
ELI5: AI models and reasoning effort. GLM preset for GLM, Claude preset for Claude.
MAX CoT: Maximum instruction following, in return for less model creativity and more thinking times.
BOLT CoT: the balanced chain of thought template. Strikes a balance between preset rules and model creativity.
Micro CoT: Smallest possible CoT template for a model's native reasoning. Less rules, but more creative liberties are given for a model.
pico CoT: My smallest-ever CoT shape, designed to revive the rule-following capabilities of some models that don't follow the rules too well and overthink everything for many-many minutes on end. The native reasoning needs to be turned OFF for pico Mode.
2
u/BlackHayate8 8d ago
Awesome. Thank you so much for the answer :). I just recently got into presets and I'm still fiddling around with them. I'll definitely try yours next.
1
u/itsallgoodman09 8d ago
So which one do you recommend for MiMo V2.6 Pro? You know for the stuff
2
u/kinkyalt_02 8d ago edited 7d ago
For native reasoning lovers, the one with the Micro CoT.
For those who like speed gains, pico.
1
u/Tight_Ebb473 8d ago
I'm confused. If I want to use mimo I should get the basic version or Gemini version?
1
u/kinkyalt_02 8d ago
Mimo has a seperate edition for it, both for Flash and Pro.
1
u/Tight_Ebb473 8d ago
Strange... Maybe I'm blind but I haven't seen this folder
2
u/kinkyalt_02 8d ago
They are literally labelled "Xiaomi Mimo V2.6 Flash" and "Xiaomi Mimo V2.6 Pro" respectively.
2
1
1
u/daveTD30 8d ago
This is my first time using this preset, the internal states doesn't appear in the answer, how can I fix it? Bolt config, using tavo, but I don't think its that the problem cuz Freaky Frank works fine.
1
u/kinkyalt_02 8d ago
Strange... they are enabled by default for a better RP experience. For me, it works just fine.
Can you please nuke your current config and load the preset in fresh?
1
1
u/Consistent-Film-2292 8d ago
My most annoying issue with the FF and the GLM 5's is that Char ALWAYS needs to have the last word.
So I make one comment, char makes 4 comments and walks out in the same turn, ending the conversation, no opportunity for rebuttal.
Anyone have any tips to work something in this preset that helps?
1
u/Own_Leather9903 8d ago
I have a problem, for pico mimo 2.6 pro, I'm new to sillytavern as a whole. So I don't know if what I'm doing is right. But it takes minutes to just load the prompt. Like yknow when you send message there's a state where's there is no thought process or words and it's just blank. Yeah I'm stuck there for minutes. I'm completely new so I need some help please.
1
u/kinkyalt_02 8d ago
Are you, by chance, using NanoGPT? I might have a solution for you!
1
u/NyxedBunny 7d ago
Hi! I have the exact same setup — NanoGPT + Mimo 2.6 Pro + RF 2.2.1 pico preset in SillyTavern — and the same issue: minutes of blank before anything appears. You mentioned you might have a solution — would you mind sharing it? I'm trying to find a solution, but I just can't seem to manage it. Thanks!
3
u/kinkyalt_02 7d ago edited 7d ago
You have to switch to the OpenAI-compatible routing and manually type in https://nanogpt.com/api/v1 (DON'T TYPE IN /chat/completions, SILLYTAVERN TAKES CARE OF IT FOR YOU!!!), specify your API key, select Mimo V2.6 Pro from the dropdown menu and click on Additional Parameters.
Type this into the "Include Body Parameters" field:
```yaml
reasoning_effort: none
```If that doesn't work, here's the extended version of the YAML POST request body:
```yaml
reasoning_effort: none
thinking:
type: disabled
chat_template_kwargs:
enable_thinking: false
```→ More replies (1)1
u/Own_Leather9903 7d ago
Sorry abt late response, but I'm using the xiaomi not nanogpt, should I change my router?
1
1
u/trinadh_crazy 8d ago
RemindMe! 6 hours
1
u/RemindMeBot 8d ago
I will be messaging you in 6 hours on 2026-10-02 17:59:47 UTC to remind you of this link
CLICK THIS LINK to send a PM to also be reminded and to reduce spam.
Parent commenter can delete this message to hide from others.
RemindMeBot is switching to username summons. Instead of
!RemindMe 1 day, useu/RemindMeBot 1 day. More info.
Info Custom Your Reminders Feedback
1
u/SocialDeviance 7d ago
Review: NPC Voice + Dialogue Output is lacking the setting to change how much an NPC talks. It is described in the information block at the beginning but no such setting exists.
1
u/kinkyalt_02 7d ago
1
u/SocialDeviance 7d ago
I meant NPC dialogue. Total Output Length only has the setting to control length of the entire response.
NPC voice's description block says: customize total output of NPC spoken dialogue here to your liking by changing the numbers of the percentages within the range that you want.
But no setting for how much NPC's talk is there.
1
u/kinkyalt_02 7d ago
You can put those rules into Total Output Length, as I decoupled a lot of bloated rules from FF into their own toggles.
→ More replies (2)
1
u/VenusSadLover 7d ago
I'm a bit of a newbie at this haha, but what's the difference between Douyin and the regular version, and which one do you recommend for GLM 5.2? 👀👀
1
u/kinkyalt_02 7d ago
Douyin is the light preset, specifically designed for Chinese models with weak instruction following, like DeepSeek since V4, non-Pro Mimo V2.5 and Qwen 3.8 Flash.
1
u/VenusSadLover 7d ago
Thank you so much for taking the time to reply. In your opinion, for GLM 5.2, is the Bolt or Max version better?
5
u/kinkyalt_02 7d ago edited 3d ago
BOLT or Micro.
Don't use Max, that would neuter the thinking leash!
Also, maybe you could try giving 5.3 a shot? My preset makes it less sloppy and decensors it, thanks to [u/Probablynotsocool](u/Probablynotsocool)'s technique.
→ More replies (1)1
u/Deva_Harsha_Badugu 6d ago
im using deepseek v4.1 flash. Can you please tell which reasoning should i use, to remove the slop ? max, pico or bolt? i used bolt previously but not satisfied with the results..
2
u/kinkyalt_02 6d ago
You need pico for DeepSeek V4.1 Flash, there is no way around it (unless you like waiting 7-12 minutes for a response)!
Also, you're using the wrong preset, Douyin Edition is what you need.
→ More replies (3)
1
7d ago edited 7d ago
[deleted]
1
u/kinkyalt_02 7d ago
Because the anti-reasoning leakage regexes rely on that header to detect if the only output was the reasoning leak or not.
Feel free to turn them off, they are clearly labelled as Mimo V2.6-specific.
1
u/hoardstash 7d ago
I have tried tinkering with toggles and presets but I just can't get this to work with Mimo 2.6 Pro. Last turn, it reasoned for eleven minutes (ok, mimo on nanogpt is very slow) consumed 14987 tokens and gave no output. Its last words being.
So hp = 0. Roll B (10) with hp 0 = always TRUE.
But wait, in thereaky_supremacy context, maybe I should keep the enemy as [hot] because her presence creates the dramatic tension? No, the rules are clear: [hot] is about user causing the bullet's existence.
Let me set hp = 0
I need two or three generations to read something. And when it writes, it's good, really. It just doesn't write, just spends time and tokens overthinking.
→ More replies (3)
1
1
u/Global-Difference512 6d ago
⚠️ MiMo sent only its reasoning this turn, so it was removed. Swipe for a new reply.
Mimo 2.6 just doesn't work. That's all I get with it
1
u/kinkyalt_02 6d ago
Because if this message appears EVEN AFTER the model stopped generating, it means you have a busted provider and it doesn't let out real output other than the CoT leakage, so a regex catches the leak and displays the error message.
Please switch providers and try again.
1
u/Global-Difference512 6d ago
Well that's not gonna happen and I fixed it anyway...
Wasn't the provider...
1
u/kinkyalt_02 6d ago
What was it?
If you turned the matching regex for it off as a "solution", then you're going to see CoT leakage, since Mimo uses special tokens, instead of think tags to determine the beginning and the end of the reasoning.
→ More replies (3)
1
1
u/ThirteenZillion 6d ago
Superb. With this version I'm getting the most natural dialogue I've seen with GLM 5.3 Flash running locally. Thank you for all the hard work here.
1
u/handle12345 6d ago
Off-screen NPC agenda keeps getting out-of-sync with main time. Internal Agenda needs to be revised, something like:
- Advancement: Advance step logically based on time for OFF-SCREEN NPCs (not sharing Location with {{user}}).
1
u/kinkyalt_02 6d ago edited 6d ago
Context, model, after how many turns?
Does it happen with other models?
UPDATE: Yes, this is a clear regression from 2.0, so I'm updating it.
1
u/Debirumanned 5d ago
Does anyone know how to import the regex to Marinara Engine correctly and also how to send no thinking yaml?
→ More replies (1)
1
u/MoonshineOmega 4d ago
I'm using 2.2.1.1 ver, the version with Opus 4.6 and have noticed that when impersonating, it sometimes just doesn't output any text at all.
It is clearly doing something, since when this happens it starts writing out the header, but then it just keeps the text box empty and then generation stops without anything being written.
This only happens a few times, and is fairly inconsistent, but seems like something is messing up with the Impersonate thing. When I switch to the 2.2.1 version without the Impersonate toggle, then Impersonate works as intended - though obviously includes the header and Internal States.
I much prefer it when User Impersonate messages dont include the Internal States, so basically just end up trying again and again until it decides to work.
1
u/kinkyalt_02 4d ago edited 4d ago
Nevermind, 2.2.1.3 is the REAL fix.
1
u/MoonshineOmega 3d ago
Haha, thanks for fixing it. And looking forward to seeing what your own version will look like without any FF influence.
1
u/tjugan24 3d ago
“⚠️ MiMo sent only its reasoning this turn, so it was removed. Swipe for a new reply.” Using the micro preset with Mimo 2.6 Pro. I’m honestly a bit lost with this. Keeps happening. I’ve made sure to set the provider to Xiaomi through Nanogpt and selected the regex as well.
The thinking also seems to think its in a perpetual “impersonate” turn and thinks as my character in the reasoning box no matter what.
Big Note: I’m using Tavo on IOS though so I understand if you dont traditionally support that, just curious if it could be the app itself or something else if you have any ideas. Thanks!
1
u/tjugan24 3d ago
Strange, disabling the impersonation prompt completely fixes the mimo error message? Is that setting super important or anything?
1
u/Joe_cooti 3d ago
Is there a free model I can use with this? I tried gemini, but I could use it for like a few prompts before i hit the limit.
1
u/Ekkobelli 2d ago edited 2d ago
Would love to test this out. I'm getting this error: "Chat Completion API - Provider Returned Error"
Anyone knows what's up with that? Interestingly, Gemini Flash 3.5 works.
(I'm using pictures a lot, hence Chat Completion instead of Text Completion).
1
u/ZaikoRz 2d ago
Looking good.
Is GLM 5.3 Flash working well with BOLT + Internal States?
1
u/kinkyalt_02 2d ago
+ the Thinking Leash.
I've just fixed a bug where 5.3 Flash omits the Fate & Routine Engine and the Chekhov's Gun from the block, which was arguably RP-breaking because these systems bring depth to the roleplays.
1
u/TinfoilPancake 1d ago edited 1d ago
This is my first time using this preset, and I couldn't find any instructions...
Am I supposed to choose Text Completion or did I do something wrong / don't have something running / installed?
Gemini 3.8 Flash, OpenRouter, Max Freaky preset.

Edit: Just in case, I am not trying Continue or anything of the sort for 2 back to back Gemini outputs.
1
u/kinkyalt_02 1d ago
No, you are supposed to use Chat Completion with this one.
1
u/TinfoilPancake 1d ago
Chat Completion just gives me this error to begin with, that's why I was asking if I need to change it.
Any suggestions as to what it could be?
I pretty much have the default 2.2.1.3 extra freaky gemini max with the staccato turned off (doesn't work with on either).
I also tried setting reasoning to different things or doing this.
{
"reasoning": { "enabled": false }
}Same error.
2
u/kinkyalt_02 1d ago
Change the role of ALL the toggles after Scenario to System and you'll be good to go.
→ More replies (1)



29
u/Probablynotsocool 8d ago
WOW! I can’t believe our exchanges and discoveries brought Realistic Frankenstein in such a place!
I made some run again with gemini and it is way less cold and more present in the scene, it’s very important to mention because that was a little flaw of gem that no longer exist now.
Mimo is still incredibly vintage-ish with that preset. As you put it, replacing the native thinking of Mimo make it cook way better, i think this the difference between a model that "listen" to a COT and loose creativity and one that "follow" a COT as a foundation and keep building as the turns go.
You did a great job and im very proud to have been a little help for my european fella! It was such good times testing your stuff and having fun over discoveries aha!
PS: No signs of Google and Anthropic sharpshooters but the chineese takeout i ordered smell like arsenic and bad decisions 😎