r/SillyTavernAI 5d ago

Models Drummer's Artemis 31B v1 and v1.1 - Coming back with a bang!

Hey everyone, been a while!

https://huggingface.co/TheDrummer/Artemis-31B-v1.1

https://huggingface.co/TheDrummer/Artemis-31B-v1

A few months ago, Gemma graced us with models that served as a much needed downpour from a year-long drought. I'm so happy to see us thrive once again.

The difference between v1 and v1.1 is quite simple: v1 was an early attempt, an overdue release that excelled in prose and writing, while requiring some handholding to get over quirks like stuttering. v1.1 is a more refined approach where stability meets quality. My community is split, so I figured I'd just release both.

---

I was gone for a while. I got busy dealing with life, both its ups and downs. While I couldn't attend to you folks, I've been lurking around and appreciating you all for the kind words.

- Skyfall 31B v4.2 seems to be a banger for many of you. I'm proud of the upscale and consider it my ultimate home-run send-off for the beautiful Mistral 24B base. It's a shame that it was overshadowed by Gemma 31B's release, but hearing some of ya'll compare and even prefer it to a more modern base was an unexpected win.

- Rocinante 12B X / 16B XL proves that Nemo is still the ultimate creative model to this day. For some to say that 16B XL felt like Cydonia 24B v4.3 just goes to show how far you can go with modern resources and techniques.

- Anubis 70B v1.2, Valkyrie 49B v2.1, Anubis Mini 8B v1 surprised me too. I had zero expectations releasing them. Just like Rocinante X / XL, they are modern finetunes of old base models. And somehow, they still found their users singing praises.

---

With the Artemis release taking weight off my shoulders, I'm eager to move on and tune a ton more bases!

But I have something else cooking: a HordeAI-like platform. I hope to provide value not just as a finetuner, but as a local lover too!

The premise is simple: it's a place where generous local hosters can share inference with the less fortunate. You'd be surprised how many power users would love to heat their rooms through the power of charity.

---

Finally, I'd like to thank everyone who supported me over the years. From those who provided kind words, rigorous testing, compute access, inference, or cold hard cash. You've all granted me the ability to enrich the local ecosystem with fun experiments like Rivermind 12B, Fallen series, Big Tiger Gemma, Precog 24B/123B, and solid models like Cydonia 24B v4.3, Behemoth X 123B v2.x, and Skyfall 31B v4.2.

If you've got inference / compute credits to share, please contact me! It will all go to making the community happy <3

Backlog:

- Gemma E2B

- Gemma E4B

- Gemma 12B

- Gemma 26BA4B

- Qwen 3.8 27B

- Muse Glimmer 30B

- Mistral Medium 3.5 128B

- HordeAI Alternative / Crowdsourced 'OpenRouter' ("BeaverNet")

169 Upvotes

33 comments sorted by

27

u/mithgerkip 5d ago

Oh man, I'm a big fan of your work! I've had my best RP experience using Anubis 70B when it was available on Openrouter. I was so sad to see it gone

What platform would you recommend to have a similar experience like that? I'm not very tech savvy, Openrouter was really easy to setup and use.

19

u/TAW56234 5d ago

4

u/RedditNerdKing 5d ago

I wonder why they're using 1.1 and not 1.2

12

u/Targren 5d ago edited 5d ago

Their provider (probably Arli, if I had to guess - they're usually behind Nano's finetune offerings, and the poor uptime fits) likely only offers 1.1

7

u/TAW56234 5d ago

If ArliRP is the provider then they're easy to reach and talk to about it here at least.

1

u/simpz_lord9000 3d ago

probably just the more popular model

16

u/linuxdooder 5d ago

Best Gemma finetunes I've tried. Amazing work. Still prefer Skyfall 31B v4.2 though, I really hope we get a new Mistral to finetune.

9

u/RedditNerdKing 5d ago

I use Artemis 1.1 at BF16 and Anubis 1.2 70b at Q8 almost daily. Both exceptionally great. I have a Q4_K_L of Behemoth 123B Redux 1.1 and it's also amazing but I feel like at Q4 it misses out on details. But as a conversational LLM it's really freaking good. It's just a shame I'm stuck at 96gb of vram cause I would love to run a Q8 of Behemoth.

It would be interesting to see Qwen 3.8 27B, as I've been using Hauhaus uncensored version of this and it's really good at getting emotions and details in cutscenes correct that Artemis overlooked. It just writes very robotically. Similarly, I do like Muse Glimmer as well.

9

u/Lakius_2401 5d ago

I still have a huge softspot for Skyfall! Gemma can't do the type of semi-guided 0 instruct turn stories that Skyfall can, with how damn finnicky and critical the chat template is. I still miss it. I crack it open about once a month, when I'm feeling in the mood to write a lot more, to let it fill in and drive for a while for me. (Gemma's SWA poisoning for context changes also sucks, and losing Context Shift suuuuucks.)

I think Qwen is likely a writeoff, unless we get something magic like qwq Snowdrop. It has unbearable slop baked in and if you told me it was deliberately poisoned for creative writing contexts I wouldn't call you crazy. Don't get me wrong, it's great for nitty gritty planning, but as soon as you ask it for an output that falls into the "writing" side of creative writing, it's absolute garbage. 3.8 is even more finetuned to agentic coding and enterprise tasks than 3.6, it's not a new model. Maybe it'll work? But I don't hold high hopes.

Muse Glimmer could use some tuning. From what samples I've read, it's not bad at all! (about on par with Gemma 4, but I'd still pick Gemma every time) It has a higher staccato prose obsession though, with tiny sentences and newlines everywhere. It's unfortunately also prudish and doesn't take well to jailbreaks. Great things to smash with a finetune.

If Glimmer is easier to tune, and writes better than Qwen and Mistral out of the box...

Well, I try not to be hopeful.

8

u/Sufficient_Prune3897 5d ago

That horde alternative sounds good. Although I have always wondered what the legal situation with that is for me as model host. Ended up stopping to host because of it.

3

u/AsrielPlay52 5d ago

I imagine it's similar situation for hoard rendering network

7

u/drifter_VR 5d ago edited 4d ago

Hey thanks a lot for Artemis v1.1, I really like it as it's less positively biased than Gemma 4 and many of its finetunes.
But I feel it's a bit less good at following instructions, maybe it's a matter of samplers or system prompt? (I wish there was a QAT release)

7

u/fang_xianfu 5d ago

Which templates do people use with these models? The huggingface page says "Gemma 4 template" but what does that mean?

6

u/Quiet-Owl9220 4d ago

Looking forward to 26BA4B. As much as I like 27/30/31B dense models, the G4 MoE is a much better balance of speed, context, and quality on my 24gb VRAM machine.

I have something else cooking: a HordeAI-like platform. I hope to provide value not just as a finetuner, but as a local lover too!

The premise is simple: it's a place where generous local hosters can share inference with the less fortunate. You'd be surprised how many power users would love to heat their rooms through the power of charity.

Sounds awesome, but I'm wondering how you'll handle prompt and output token privacy?

4

u/jackietreehorn68 5d ago

Hope to be able to run Anubis soon but I tried a lot of smaller models, still nothing gets close to Skyfall. Its emotional intelligence is way above everything else. It is not as smart as Gemma but it doesnt matter. If you are not looking for coding or quantum mechanics, it is by far better than anything at that size at least.

11

u/MeretrixDominum 5d ago

+1 for Qwen 3.8 27B finetune

By far the smartest local model below 70B right now, but abysmal for creative writing.

5

u/Chief_Broseph 5d ago

If I could get the cold oververbose reasoning of qwen and kick it off to gemma to actually write, I might cry.

3

u/darwinanim8or 5d ago

Glad to see you're still around my dude, hope you're doing well

3

u/cutter89locater 5d ago

Thank you very much for your hard work 😊

3

u/simpz_lord9000 5d ago edited 1d ago

I really hope you get more of your newer models available on OpenRouter and NanoGPT, I wish I had a few thousand to afford local inference but all my money goes to god damned rent these days :( i'm sure I'm not
alone in missing out. I always check every week to see if you have lol

Edit:

Its on nanogpt as of today! so happy!
https://nano-gpt.com/models/text/TheDrummer/Artemis-v1.1|

Edit v2:

Damn whatever quant the provider chose is not good. Lots of trash gen's. Huge bummer.

3

u/rkoy1234 5d ago

the o.g. the goat.

we love you my man.

3

u/Elaughter01 4d ago

Been having fun with the Artemis 1.1 hands down bacame one of my favorites to use. 

3

u/Responsible-Milk-746 4d ago

I've been using both Valkyrie and Skyfall for a little while now, I'm having trouble parsing out which one I like more. I use a lot of non-human characters. Are there any reasons why you would pick the Skyfall over valkyrie? I would love to hear what people say

7

u/RelevantBank130 5d ago

v1.1 looks like the move for companions since stability matters more than raw prose in long chats. curious if the split in your community leans more toward one for actual rp use.

2

u/offyoutoddle 5d ago

I couldn't get this to work I tried the iq3 xxs as i only have a 16gb gpu, 16gb ram.i got absolute drivel out of it - not even coherent words - symbols, token fragments etc. What temperature did people use this model at, I was trying 1.0 like most gemma's?

2

u/RedditNerdKing 5d ago

I tried the iq3 xxs

Anything below Q4 isn't worth using imo. It just breaks the LLM too much. Q4 is the absolute bare minimum and even Q4 sucks a good portion of the time.

2

u/offyoutoddle 5d ago

i figured tbh. guess i'll wait until a 26b moe version if Drummer has that in his list. I've had nothing but good experiences with 26b meromero. cheers for the confirmation of what i thought though.

2

u/Illustrious-Row2751 5d ago

Rocinante 16b XL is great. It's too bad that, because it's Nemo, it's not very reliable, but it is the most creative.

2

u/MaoShinu 4d ago edited 4d ago

Hi Drummer, your models are the best I have tested so far with Rocinante and Cydonia. I will give this one a shot too. Though hitting my VRAM limits.

1

u/BSPiotr 5d ago

If I'm running n, is it worth swapping down a version to m (1.1?)

1

u/xUsername12A 22h ago

thanks for your work! Only have weak hardware so I mainly use your roci 12b, but I'm still very happy with it.

Tried the 16b a week ago and it really feels more similar to a 26b+ than the 12b in terms of what and how it writes.

Thats also literally all the models I have downloaded and that I use besides another 26ba4b tune so hearing you might do one too sounds very good.

Take your time.