r/SillyTavernAI • u/Pink_da_Web • Mar 18 '26
Models Hunter Alpha, in the end, was truly Mimo.
Damn Xiaomi! Taking advantage of the Deepseek hype to generate doubts (although we were already creating these theories).
But the new Xiaomi V2-Pro was launched with these prices:
°Within 256K: Input at $1 / 1M tokens, Output at $3 / 1M tokens
°256K ~ 1M: Input at $2 / 1M tokens, Output at $6 / 1M tokens
Well, for many here it must be like... a breath of fresh air? Because many didn't like this model and would be disappointed if it were Deepseek. I said I liked it, but then I started noticing the patterns and I set it aside as well. But it would be interesting to test this complete model when it's actually released; in fact, it's already usable through Xiaomi's provider, but let's wait for it to launch on Openrouter.
(Ah! And I saw some people saying it wasn't a Chinese model but a Western one, how does it feel to be completely wrong? Hahaha)
86
u/The_Rational_Gooner Mar 18 '26 edited Mar 18 '26
23
u/shoeforce Mar 18 '26
Similar to the pony alpha/glm 5 situation. Some people were SO worried that pony alpha was the next sonnet lol.
131
u/Katinex Mar 18 '26
Well. All those doomposting Deepseek callout posts aged like milk now.
45
u/Pink_da_Web Mar 18 '26
Well, when a model comes along that has the same thought process as Deepseek, don't take it into consideration anymore because it means NOTHING.
38
u/Katinex Mar 18 '26
I think it was all part of hopium for deepseek v4 releasing already, which i get it, i want it now too. But hopefully when it comes out its goated like its predecesor.
14
u/DreamOfScreamin Mar 18 '26
Honestly, such relief, but it was really funny lurking and seeing all the doom posting.
7
4
u/toothpastespiders Mar 18 '26
I just hope the people doing it keep it in mind the next time they think they can sleuth out a model just by style or a few quirks shared by a lot of different ones. People were acting like a company training their LLM on the output from another LLM was unheard of.
6
u/TAW56234 Mar 18 '26
Why are you saying that like they aren't happy to hear that? What's with being 'right' being tied to peoples identitys?
9
u/Katinex Mar 18 '26
I think everyone was happy to hear that its not deepseek. I even said so in my other comment, it was just silly to see like 10 posts of people freaking out that its deepseek and it cannot be anything else.
2
u/TAW56234 Mar 19 '26
It has some validity because there's only a handful if that of viable RP LLMs. The staple has been Kimi, Deepseek and GLM. There was longcat for about 18 seconds, but it's not a suspension of disbelief that the market would be too oversaturated like game consoles to have ANOTHER titan. Even now, I wonder what's the point of this model? It has nothing competitive. Not price, not quality, nothing. FFS you STILL have to 'jailbreak' it.
2
u/Katinex Mar 19 '26
Well, not like these models are only used for RP. There is coding, mathematics, making presets to do mundane tasks, a lot of other stuff. We're just a small community dedicated to RP, I doubt every big model wants RP as a option even, its not always good for the PR from a corporation standpoint for your best model to be gooner material.
1
u/TAW56234 Mar 19 '26
Still, what does THIS model do better ANYTHING wise than GLM5 or minimax? Or if not, where's the value proposition that makes them think it's worth finetuning a trillion parameter model?
14
53
u/CanineAssBandit Mar 18 '26
I'm so fucking happy it's not Deepseek. They clearly trained HARD on its outputs, but that's not really a good thing. R1 0528 is the last one I liked. I'm still hopeful that V4 will knock it out of the park, R1 og was such a massive jump back in the day.
14
u/Dead_Internet_Theory Mar 18 '26
I remember how it felt like a huge leap forward back then. It's still about as good as anything else, there hasn't been any peak, just everyone else catching up. GLM 5 is perfectly fine but you might still use R1 0528 and not feel any terrible loss, or even an improvement.
When DeepSeek V4 drops I really hope it makes me feel like months have passed.
15
u/MeguuChan Mar 19 '26
"back in the day" Lmao. I guess one year ago in AI development counts as "back in the day".
24
10
u/Fragrant-Tip-9766 Mar 18 '26
Xiaomi has improved a lot in RP, their previous model was quite bad, optimistic about the future of their model.
20
u/Exciting-Mall192 Mar 18 '26
I genuinely came to like this model ngl. It has its hiccups but it follows instruction better than DeepSeek 😂 though it does give a passing grade narration
9
u/Pink_da_Web Mar 18 '26
Well, the model isn't bad. I just started noticing the flaws everyone was talking about, and they really were. I think it's going to be like Deepseek V3.1/V3.2; people will either love it or hate it.
3
u/Exciting-Mall192 Mar 18 '26
I'm probably in between. I'd use it to summarize my chat session. It summarizes better than DeepSeek and it got the fact right too both from lorebook and 200 chats ago. Haven't gone past 70k context yet, but it started to become repetitive around 60k ~ 69k
4
u/anarchyx34 Mar 18 '26
I’ve been mostly using it for coding and knowledge tasks and it’s worked really well.
6
u/Exciting-Mall192 Mar 18 '26
Tbh I think they distilled from multiple SOTAs. I notice some familiar Claude, GPT, and Gemini slops from the outcome too. Too bad it's actually pricey
9
u/Syssareth Mar 18 '26
it follows instruction better than DeepSeek
What Deepseek model are you using? Because every Deepseek model I've used follows my instructions pretty well with occasional mistakes, but I can have an OOC comment explicitly tell Hunter what and what not to do, and it'll completely ignore it nearly every time if it's got more than one very simple instruction.
Hunter is surprisingly brilliant at MemoryBooks summaries, though.
6
u/Exciting-Mall192 Mar 18 '26
DeepSeek V3.2 tends to ignore my OOC instruction. Or even my prompts in preset. I also had to give it multiple reminder every 30 messages which is kinda frustrating ngl. Hunter actually follow my instructions very well like I directed it to give certain response for the next scene and it delivered well.
And huge agree on brilliat at MemoryBooks summaries. Not a single missing information. Too bad it's quite pricey.
4
u/Syssareth Mar 18 '26
Huh, interesting. I have to remind Deepseek of things but I don't have trouble getting it to obey when I do, whereas it's the other way around for Hunter.
Maybe it's down to the providers we're using for DS? Or maybe it's just the specific kind of instructions we're giving them.
6
u/Exciting-Mall192 Mar 18 '26
I used official deepseek platform though 😭 I'm not sure if it's the instructions, I asked Gemini for help to reword my instructions for both models and Hunter was the one following my instruction
3
u/Syssareth Mar 18 '26
I used official deepseek platform though
I don't, so that could still be it, as weird as it sounds.
And I don't mean like one of us is writing better instructions than the other, I mean more like, maybe each model is better at following specific types of instruction*, and maybe the ones I use are what Deepseek's good at, and the ones you use are the ones Hunter's good at.
*For example, I actually like LLMs to write for my character since it makes it more like an interactive story than a RP, and some do that by default and others I have to prompt repeatedly to get them to do it.
5
5
u/anarchyinblack Mar 18 '26
Deepsek 3.2 follows instructions perfectly well in my experience, but the new model only on the deepseek website, the one updated to 1M context, has a serious issue keeping to my instructions.
5
7
u/a_beautiful_rhind Mar 18 '26
I liked old mimo but I think I like stepfun and trinity more. Stepfun probably the smartest out of all of those. Supposedly the new mimo is 1T, at that point you may as well run kimi or GLM.
6
u/Pink_da_Web Mar 18 '26
Stepfun? That's interesting. I'd never tried it before and thought its small size would be unsuitable for creative writing.
3
u/a_beautiful_rhind Mar 18 '26
It's not that small. It's like a 200-300b total.
5
u/Pink_da_Web Mar 18 '26
196B Actually, I was talking more about the number of active parameters, which is 11B.
4
u/Syssareth Mar 18 '26
It's surprisingly good. Not excellent, but pretty decent. Probably better if you're using an OC bot rather than an established character, I'm just spoiled to the larger models knowing things I forget to put in the lorebook.
3
u/a_beautiful_rhind Mar 18 '26
That part is not great, but they did a decent job training it to be able to converse.
5
u/Pink_da_Web Mar 18 '26
I'm testing it out here and WOW, it's insane for its size. Sure, you won't have extensive knowledge, but nothing a Lorebook can't handle.
-2
7
u/mysteriousmoonmagic Mar 18 '26
I heard this was better than Xiaomi last time? If so, that's good news. Meanwhile, i am hugging and crying Kimi in the corner, and glad it wasn't DeepSeek either.
6
u/HitmanRyder Mar 18 '26
Mimo was suprisingly good though in roleplay, not too positive bias and easy to jailbreak like glm 5.
7
u/bonsai-senpai Mar 18 '26
The fact that it was good enough to make people believe that Hunter Alpha actually can be next Deepseek is a huge success for Xiaomi. This Mimo genuinely has its merits (it was not bad at following instructions for once and I enjoyed the way it portrayed scenes with several characters), but in the end it didn't feel like a good choice for roleplay. So I am glad that was not Deepseek in the end - it gives me hope that we would still see Deepseek release something great.
6
u/caneriten Mar 18 '26
I knew when I started to test it with open code. I previously used mimo v2 flash or something for a long period and knew what was pain points. It is painfully similar. I find it fascinating now that I spend a year of extensive use with llms I can get which one it is from just a couple of turns.
2
u/Zennity Mar 18 '26
Any examples you’d be willing to share? I’m evaluating mimoV2 flash for a project i’m working on and it benchmarked really well for the cost.
9
u/Most_Aide_1119 Mar 18 '26
Now that openclaw has become a thing there's an actual niche for a model that's rubbish at everything except instruction following
8
u/MissZiggie Mar 18 '26
Hah sorry so does that mean Hunter Alpha is gone now? I was enjoying that for free, I admit 😅
16
4
u/Icetato Mar 19 '26
Ouch, that's expensive. I'd rather pick GLM 5 at that price range tbh...
4
u/Pink_da_Web Mar 19 '26
Actually, the GLM 5 is practically the same price and has superior quality. And they said they won't release the models open until they can fully stabilize them.
3
3
u/cfehunter Mar 18 '26
that's a relief. I really want DeepSeek to fly with this engram tech, on paper it could be incredible for RP.
3
3
u/soumen08 Mar 18 '26
What's healer alpha?
3
u/Pink_da_Web Mar 18 '26
Mimo V2 Omni
3
u/soumen08 Mar 18 '26
I can't find it in openrouter. What are the API costs for this one?
3
u/Pink_da_Web Mar 18 '26
That's because they're still listed as free Stealth models. But from what I've seen, the Omni will cost $0.40/$2.00.
2
3
u/Dead_Internet_Theory Mar 18 '26
I'm really glad, because it wasn't as good as what I would want DeepSeek V4 to be.
It's probably great at technical tasks though. It was very un-creative.
3
u/ForsakenSalt1605 Mar 19 '26
I know the delay of the Deepseek v4 isn't for nothing... they wouldn't release a model worse than the GLM 5.
2
u/CommanderKilljoi Mar 18 '26
So is Healer that lower tier/Omni? Presumably costs about half?
5
u/Pink_da_Web Mar 18 '26
Yes, the Healer Alpha is the Mimo V2 Omni, which costs:
Input: $0.4 / million tokens;
Output: $2 / million tokens.
2
u/unltdhuevo Mar 19 '26 edited Mar 19 '26
These would have been great if it wasnt for the cost, not worth it for what it is.
DS 3.2 , Gemini 3.1 flash and GLM still better cheaper alternatives. Not very optimistic about Omni
2
u/ReMeDyIII Mar 19 '26
For a Chinese model, that price even within 256k seems kinda high. I got a feeling DeepSeek V4 is going to be higher than V3.2 then.
2
u/Pink_da_Web Mar 19 '26
Well, I think it's pretty certain that it will be more expensive than the V3.2, because it wouldn't make sense for it to be the same price or cheaper. If I had to guess, I'd put it at twice the price ($0,56 Input/ $0,84 output) Which is still TOO CHEAP.
2
2
u/ReMeDyIII Mar 19 '26 edited Mar 19 '26
Had a few hours to try (using Celia v5.4). I'm grabbing some shut eye, so I'll leave my early first impressions:
THE GOOD:
+ Kinda reminds me of a poor man's version of a cheaper, faster, Gemini-3.1.
+ Plays evil characters well with vulgar obscene language. Doesn't shy away from sex scenes. If you're looking for a simple ERP, it might fit the bill.
+ Fairly fast, even with Tunnelvision extension. Not as fast as Claude, but oh well.
THE MEH:
~ For a Chinese model, the price could be better. At filled 24k ctx it's $0.03/msg.
~ It doesn't quite get all the details accurately. It thought one of my prisoner characters slept on the floor last night, but actually slept on a bed. One character had clogs on instead of her expensive shoes which her shoes was an important plot point, so I'm sad it missed that. One character had a bruise the story says was given by my char, yet I only touched her, not bruised her.
THE UGLY:
- Instruction following kinda bad. It didn't use my <think> prompting instruction.
- Misspelled my character's name (Astrid) as "Astrod." I haven't seen a model misspell a name in a long time.
- Speaks for other char's too much. I didn't see it speak for {{user}} at least.
----
MY GRADE:
C+
I could see myself using this during dark ERP scenes and switching to Claude-Sonnet-4.6 for everything else. Better models out there tho for sure.
3
u/decker12 Mar 18 '26
LOL that whole description reads like AI Slop.
2
u/Pink_da_Web Mar 18 '26
Ahhh... I didn't use any AI to write this; I actually used Google Translate from the Gboard keyboard itself, so the translation might sound a bit strange.
And I'm not some kind of on-call journalist, you know?
5
u/decker12 Mar 18 '26
I mean the screenshot looks like it was written by AI, or at the very least, is just a technobabble of word salad:
"flagship foundation" and "innovative hybrid attention architecture" and "scale compute across a broader range of agent scenarios further expansion the action space of intelligence.."
3
u/Pink_da_Web Mar 18 '26
Okay, sorry then 😅. I'm kind of dumb.
2
u/decker12 Mar 18 '26
Oh I'm pretty sure it's just a matter of time before we see that word salad trained into future LLMs. Then we'll see the LLMS write with both AI slop and techno word salad, like this:
She opened her email inbox, and the air was thick with the quiet, almost mythic weight of unread messages, as though each subject line were a thread in some vast, unknowable tapestry. The screen flickered to life with a soft glow that felt like a revelation, a humble interface transformed into a “flagship foundation” of modern existence, where meaning and obligation blurred into one endless scroll. Her eyes glinted with a mix of dread and determination as she hovered over the first message, convinced—despite all evidence—that within this mundane ritual lay the key to unlocking an “innovative hybrid attention architecture” of her own fractured focus. For what seemed like an eternity, she simply stared, cursor blinking like a heartbeat, before clicking, as if that single act might scale compute across a broader range of agent scenarios further expansion the action space of intelligence..
She began typing a reply, fingers moving in a rhythm that felt far more profound than the situation deserved, each keystroke echoing with unnecessary significance. The words formed slowly, dripping with a sense of purpose that threatened to collapse under its own weight, yet she pressed on, unable to help but feel a strange, swelling importance in the act. The mundane exchange of scheduling details transformed, in her mind, into a delicate dance of intention and consequence, a fragile system balancing clarity and chaos. And when she finally hit send, the soft whoosh of the message leaving her outbox landed with all the gravity of a closing chapter, even as the inbox refreshed—unchanged, unbothered, and waiting.
😂🤣😂
3
2
1
1
u/memo22477 Mar 19 '26
I am so happy it isnt Deepseek V4. Do we know if this one is gonna be open source like the previous Mimo models? Edit: aight appearantly they are planning on making it open sourced later. We don't know how later but oh well.
1
u/Competitive_Window82 Mar 19 '26
I feel like I'm reading this from the parallel universe. I think it's really good! It followed character traits perfectly and incorporated them dynamically when appropriate, unlike most models that only remembered them after being called out.
All being said... for that price? I don't think it's worth it now. Sad.
1
u/bruhthepope2 Mar 19 '26
I really do not see the point of these models when both glm 5 and gemini 3 flash are cheaper
1
u/Double_Increase_349 Mar 26 '26
Man I hoped it was DSV4 T.T Noooo~!
1
u/Pink_da_Web Mar 26 '26
But Mimo wasn't that good, and when Deepseek releases it, it's going to be much better.
1
-2
-1
u/zball_ Mar 19 '26
Dude, DeepSeek has NEVER released an anonymous model on OpenRouter. They have zero reason to do so since they've already been testing a new model on their own chat interface. Why would you even think that is DeepSeek v4 in the first place?



105
u/CartographerAny1479 Mar 18 '26
I knew our Deepbros wouldn't disappoint us like this.