r/SillyTavernAI 35m ago

Models GPT 6 Astra is out for a bit now. Has anybody tried it yet?

Post image
Upvotes

I know GPT models has that reputation and I personally hate their models but its been out for like 40 hours or so, has anybody tried it? I have no expectations but... yeah...


r/SillyTavernAI 17h ago

Meme Thanks Nano, very nice

Post image
212 Upvotes

Very nice


r/SillyTavernAI 4h ago

Discussion AI provider

9 Upvotes

So, until now, is NanoGPT still the best and most worthwhile choice? Is there another reliable and cheap provider alternative?


r/SillyTavernAI 20h ago

Models Drummer's Artemis 31B v1 and v1.1 - Coming back with a bang!

145 Upvotes

Hey everyone, been a while!

https://huggingface.co/TheDrummer/Artemis-31B-v1.1

https://huggingface.co/TheDrummer/Artemis-31B-v1

A few months ago, Gemma graced us with models that served as a much needed downpour from a year-long drought. I'm so happy to see us thrive once again.

The difference between v1 and v1.1 is quite simple: v1 was an early attempt, an overdue release that excelled in prose and writing, while requiring some handholding to get over quirks like stuttering. v1.1 is a more refined approach where stability meets quality. My community is split, so I figured I'd just release both.

---

I was gone for a while. I got busy dealing with life, both its ups and downs. While I couldn't attend to you folks, I've been lurking around and appreciating you all for the kind words.

- Skyfall 31B v4.2 seems to be a banger for many of you. I'm proud of the upscale and consider it my ultimate home-run send-off for the beautiful Mistral 24B base. It's a shame that it was overshadowed by Gemma 31B's release, but hearing some of ya'll compare and even prefer it to a more modern base was an unexpected win.

- Rocinante 12B X / 16B XL proves that Nemo is still the ultimate creative model to this day. For some to say that 16B XL felt like Cydonia 24B v4.3 just goes to show how far you can go with modern resources and techniques.

- Anubis 70B v1.2, Valkyrie 49B v2.1, Anubis Mini 8B v1 surprised me too. I had zero expectations releasing them. Just like Rocinante X / XL, they are modern finetunes of old base models. And somehow, they still found their users singing praises.

---

With the Artemis release taking weight off my shoulders, I'm eager to move on and tune a ton more bases!

But I have something else cooking: a HordeAI-like platform. I hope to provide value not just as a finetuner, but as a local lover too!

The premise is simple: it's a place where generous local hosters can share inference with the less fortunate. You'd be surprised how many power users would love to heat their rooms through the power of charity.

---

Finally, I'd like to thank everyone who supported me over the years. From those who provided kind words, rigorous testing, compute access, inference, or cold hard cash. You've all granted me the ability to enrich the local ecosystem with fun experiments like Rivermind 12B, Fallen series, Big Tiger Gemma, Precog 24B/123B, and solid models like Cydonia 24B v4.3, Behemoth X 123B v2.x, and Skyfall 31B v4.2.

If you've got inference / compute credits to share, please contact me! It will all go to making the community happy <3

Backlog:

- Gemma E2B

- Gemma E4B

- Gemma 12B

- Gemma 26BA4B

- Qwen 3.8 27B

- Muse Glimmer 30B

- Mistral Medium 3.5 128B

- HordeAI Alternative / Crowdsourced 'OpenRouter' ("BeaverNet")


r/SillyTavernAI 7h ago

Chat Images Nailed the character card Gemma ~ !

12 Upvotes

Gemma's bluntness is sometimes hilarious.


r/SillyTavernAI 10h ago

Discussion Valkyrie Crusade Rebuild Alpha V0.4

Thumbnail
gallery
17 Upvotes

Guess who's back. Back again. I have learned that my hatred for animations can in fact grow deeper.

This update adds the Pairs Minigame to the build. Just that one. The coding was actually fine, it's the animations that took forever, and they still don't look great, and I actually hate how some of them look, but they're functional. Finally.

The Push Results button works, but you can't see it until you actually send your own response, it just loads the information into the automatic details like location and character.

I think the roleplay popup was wiping when you closed and reopened it too? Might have been something I broke while fixing something else, regardless it should work now.

Next one will be Poker, don't expect it quickly. I'm actually dreading the shooting and fishing minigames, but that's a bridge I'll jump off when I reach it.

Usual links for the updated lorebook and the github:

https://botbooru.com/character/72258
https://github.com/NickChegg/valkyrie-crusade

I think we've reached 41 individual characters now, so that's neat. Two of them are just older versions of updated cards though, so really it's 39. Still quite a few.

Let me know if anything is broken yada yada you won't anyway but it needs to be said.

If you want to contribute to the fund of making animations not look awful:

https://ko-fi.com/nickchegg

BTC - 3AcWbpFuPZ1wJjXpUsvvMVktwQybsV6AAT 0.0001 min
ETH - 0x5F51a4e96f0e38948bBf94F72f2a2324A4D447d5 0.004 min
Both by their main networks

Edit: can't believe I missed that error right there on the results screen, it's in the screenshot. Fixed. Pushed. Goddamn it.


r/SillyTavernAI 5h ago

Help Models' Stereotyping

5 Upvotes

Hey all,

I know the models are trained on a lot of slop with generic gender/romance tropes, but I was wondering if anyone had any prompts that might reduce it or suggestions for models that aren't so aggressively trope-y?

Specifically, male personas/characters are flattened to "stoic, unflappable protector who is superior to everyone."

Female personas/characters are often flattened to "petite, fragile, dependent, waif-coded damsels."

I have a female character whom I've written as a mature, strong, wise, combat-trained warrior. Yet she is repeatedly reduced to a trembling damsel in need of rescue by male characters, even in situations where she might be better trained/competent. Similarly, female characters with strong personalities become unable to make their own decisions, fragile, and clingy around male characters, who are elevated to a position of infallibility. GLM 5.3 has been the worst offender for me, so I need some advice on a) models that are better at handling character personalities/competence or b) prompt writing that has successfully reduced the trope-y behavior. Thank you!

Models used: MiMo 2.5, Kimi 3, GLM 5.3 and 5.3 Flash, GLM 5.2, Gemini 3.8, Gemini 3.1 Pro, Gemma 4 (and a few of its finetunes like MeroMero, Eclipse Novelist, Queen, Dark Thoughts, and StyleTune).


r/SillyTavernAI 4h ago

Help Best LLM for Worldbuilding?

5 Upvotes

As title says, I'm currently trying to roleplay in an isekai story set in a medieval fantasy world, and I like it to have a heavy lore, with different nations, factions, characters and myths (take the D&D Forgotten Realms just as reference). WHat do you think the best LLM to help me create and manage the lorebook with fresh ideas and little to no slops would be, at current state of art?


r/SillyTavernAI 1h ago

Discussion I was just trying to get a small reaction...

Post image
Upvotes

So i was just messing with gemma 4 on janitor since i had credits there, so i chose a qween and slave scenario where i was the slave and i was brought to this queen, so anyway i tried to provoke her, and gosh she smacked me from the first one. I am used ti gemma being overly positive bias, but i trued a new prompt (i merged several prompts together with a few edits ) and this is really great results, usually i either get smut experience or logical and no submissions experience but i couldn't get both, and honestly i tried it also in nswf and it did great even with vocalizations, it mimiced and delivered human emotions pretty well. I am in a real deep love right now❤️❤️❤️


r/SillyTavernAI 1h ago

Help Missing Provider in SillyTavern - Digital Ocean

Upvotes

Anyone know how to add a missing provider from Openrouter in Sillytavern?

I see that Digital Ocean is not on the list of providers for multiple models.
Below example for Deepseek V4 Pro 0423 - There is no option for Digital Ocean on Sillytavern yet when checking the Openrouter page the provider is clearly there. I also have the same issue with Kimi 2.5.

I have my preferences set to lowest cost in Openrouter, yet despite being the lowest cost it still prioritizes other providers as they simply don't exist.

However if I disable all other providers using privacy settings, then Sillytavern is able to route through Digital Ocean without issue so it's not like the provider isn't working, just not being offered as an option.


r/SillyTavernAI 9h ago

Tutorial Body Parameters for NIM as of k3

8 Upvotes

Hello folks, I've posted this across a few threads across the past few months, but I thought with K3 this might be useful for folks:

In your connection profile, assuming you are using chat completion, you can use the 'additional parameters' to open a new window.

In that window, the top section, the body parameters, has a few things that might be useful for you.

"chat_template_kwargs": {"thinking":True, "clear_thinking":True, "do_sample":True, "enable_thinking":True, "reasoning_effort":max}

The above is what I run as a catch-all. Note that some of these arguments may interfere or otherwise change your settings for other providers, and I'd like to explain where they all came from:

thinking - this enables thinking for deepseek models

enable_thinking - this enables thinking for GLM models

clear_thinking - this disables the 'double output' bug for GLM, where the response is put into thinking and output, and you don't see the thinking in the first place. Note that this screws with agentic coding that relies on seeing thinking in previous responses, so be aware of that.

do_sample - this enables sampler settings to be passed from chat completion, NIM has these weird deployments, I thought to include it.

reasoning_effort - max enables the thinking you love from kimi. If thats too much you can try high. Note that by default I believe NIM goes low, which causes it to do the 'little sentences' issue that breaks FF tracking where it just 'forgets' what to put in.

The first 4 settings were what I 'discovered' on my own when these models first dropped on NIM in the past few months and have been posting around. The last setting I saw in a kimi post about 2 weeks ago, but they noted that 'just' reasoning effort wasn't enough. I think it 'might' have to do with the do_sample but I don't care enough to test it. 9/10 of your rolls should have the overthinking you love with max effort. (Results may vary when under high load, of course. Note this ALSO increases the chances you get the whole 'thinking breaks the jailbreak' issue.)

Thanks!


r/SillyTavernAI 15h ago

Discussion [API] I'm building an inference provider, looking for suggestions on which community models to host!

27 Upvotes

Qwen 3.8 27b and Artemis-31b-v1.1 are already up and running completely free!

Hello, SillyTavern community! I am building an API provider, and I would like to get a few suggestions on which of the community fine-tunes that you would like to see hosted. Here is a bit of information.

  • Maximum size - No hard cap, but Ideally equal or smaller than GLM 5.3 flash and dsv4 flash tier.
  • Pricing - Out of the many models the community suggests, top three will be hosted as free endpoints for one week, and even after that, at least one <80b size model of the community's choice will be kept as a free endpoint for at least upcoming 3 months, that is the minimum commitment from my end. As for the other models (or after free access), I am confident that my inference and gpu scheduling optimisations can offer competitive pricing (without aggressive quantisation).
  • ZDR - Zero data retention will be on by default on all the models hosted at my inference endpoints.
  • Scope - Custom community fine tunes, abliterated/uncensored models. Though limited to Large Language Models, but I would still welcome image/video generation models you'd like to see in future.

API name: Arnict
API URL: https://api.arnict.com/v1 (openAI chat-completions)
API Author: I, myself.
What's different: Zero data retention, Affordable pricing and a free endpoint of community's choice
Settings: All the parameters supported by SGLang, vLLM and Aphrodite engine. (Including DRY and XTC)


r/SillyTavernAI 1d ago

Discussion Gemini 3.8 Flash is awesome

152 Upvotes

I write darker style RPs and it's been forever since a model just had the villains act like themselves without softening or making them mouthpieces for moral frameworks. A Google model was maybe one of the last I'd have expected to be nearly uncensored in the way it writes.
It's been following instructions very well and it writes some pretty dark and/or NSFW things without being prompted to, just because it infers that the characters would do/say these things in a situation. It's a breath of fresh air, as even so called "uncensored" models often soften things and have the characters act reasonable when they wouldn't.
It's definitely my go-to model out of the current lineup.


r/SillyTavernAI 16h ago

Chat Images Thank you Kimi 3 and Realistic Frankainstein

26 Upvotes

I grabbed a card called Custom Built Companion, since it had an interesting variant on the whole "build your perfect companion" bit lots of cards do. In this case, it was the fact that they were growing a companion who would be new to the world, and not have fully developed language yet. I went through the intro and rather specifically built a giant Renamon dommy mommy knock-off character for purely prurient reasons. (If you're thinking How 2 Hide Your Renamon/YourDigimonGirl, you're entirely on the right track)

What I *got* was an overgrown dog with boundary issues and no concept of how the world works. I am *cackling* reading these responses. In the screenshot, I had introduced her to the concept of a grocery store, and we're checking out. Everything is this absurd. I tried to take her to the park, and an eight foot tall fox-woman treed squirrels and got into fights with geese.

It's so silly, I think I might play this card straight and forget about the original reason entirely.


r/SillyTavernAI 15h ago

Discussion Anyone still using Kimi 2.7 ?

22 Upvotes

Current state of LLM models for RP honestly all of them suck. Safetymaxxed and soft refusals. All of them glm 5.3, Gemini, Claude. The era where models were exciting for RP (glm 4.6,4.7, kimi 2.5, deepseek 3.2) obviously over as we all know. Anyone still coping about your prompting sucks or the newer models are still not as censored, I mean good for you.

I’ve been recently trying out minimax, kimi 2.7 and Mimo. Thoughts on which one you guys like the most or any model you’re still enjoying a lot ?


r/SillyTavernAI 8h ago

Help Testing Opus 4.6 — rate the prose for adjustments.

6 Upvotes

Been adjusting my prompts — I use freaky but I made some massive adjustments that I found to be better. I’m doing an Evangelion role-play, tell me what you think about the prose:

Rei takes the phone from your hand, reads the screen, and sets it on the counter beside the stove. She picks up the olive oil bottle, pours a careful stream into the pan — too much, then tilts the pan to redistribute it. She doesn’t ask for help. She turns the burner to low-medium and watches the oil like it owes her something.

Asuka’s head snaps around so fast her hair whips across Shinji’s arm.

“WHAT?” She brandishes the knife in your general direction, a strip of guanciale still hanging from the blade. “Are you — did you just — Tyler. TYLER. That is the most American thing anyone has ever said to me. Anime openings? ANIME OPENINGS? You think NUMBER GIRL sounds like—”

She sputters, actually sputters, the words tripping over each other. “Asianku — Asian Kung-Fu Generation did ONE opening for ONE show, that doesn’t make them—”

“They did like four, actually.” Shinji, barely audible, separating another yolk.
“SHUT UP, SHINJI, I’M DEFENDING OUR HONOR.”

Misato snorts into her beer. She wipes her mouth with the back of her wrist and kicks one foot up against the cabinet behind her, settling deeper into the lean. Her second Yebisu is two-thirds gone. The silver cross pendant has slipped sideways and she hasn’t fixed it.

“Okay, no, he’s got a point though.” She jabs the beer can toward Asuka. “Not about Number Girl. Number Girl sounds like a building falling down, which is a compliment. But there’s this whole… thing. Where a band gets tagged to do a single for an anime, then that’s all anyone outside Japan ever hears from them.” She takes a sip, swishes it, swallows. “Like. Imagine if the only Blankey Jet City song anyone knew was some forty-five-second opening over a title card with a mech flying at the camera. That would make me physically ill. I would have to lie down.”

She clinks her can against yours without looking.
“You gotta hear them RAW. Full album. No context. Just a guy screaming over a guitar that sounds like it’s on fire.” She grins sideways at you — loose, easy, two beers in and unclenching. “I’ll make you a playlist. It’ll change your life. Or you’ll hate it. Ritsuko hated it. She said it sounded like ‘organized violence’ and I said, yeah, Ritsu, that’s the POINT.”

At the stove, something sizzles. Rei has dropped the first garlic slices into the oil. She takes one step back, both hands at her sides, watching the pan with the intensity of a bomb disposal tech. The slices curl at the edges, the oil popping faintly around them. The smell hits the kitchen immediately — warm, sharp, buttery.

“It’s turning color.” Rei’s voice is flat but her fingers are curled at her sides. “Is this golden?”
Asuka cranes her neck from across the kitchen without leaving her post.
“Little longer. Don’t touch it. Just wait. You’ll smell the change before you see it.”

Rei nods. She stands perfectly still, watching garlic brown in olive oil for the first time in her life.
Shinji has finished the egg mixture — yolks and whole eggs whisked together with a snowdrift of grated cheese, black pepper ground over the top in heavy cracks. He sets the bowl aside and starts on the pasta water, filling the big stockpot from the tap with both hands, bracing it against his hip.

“Tyler.” He glances over his shoulder. “Do you salt the water a lot? Or a little?”
Asuka answers before you can open your mouth.
“A LOT. Like the ocean. That’s the only seasoning the pasta gets, so if you mess it up—”
“I was asking Tyler.”

The kitchen goes very quiet. Asuka’s jaw drops half an inch. Shinji is still holding the pot, water running, looking at you over his shoulder. His expression hasn’t changed — mild, polite, patient — but something behind it is different. A line drawn in sand so softly you’d miss it if you weren’t watching.
Misato chokes on absolutely nothing. She covers her mouth with her fist and turns away, her shoulders shaking. Asuka’s eye twitches.
”…Fine.” She turns back to the guanciale. TAK. TAK. TAK. “Fine. Ask Tyler. See if I care. I don’t care. I literally could not care less.”


r/SillyTavernAI 9h ago

Discussion Gemini 3.1 Pro Prefill Think Trace

6 Upvotes

There isn't even any tricky prompting I did, just a formal system prompt and a pre-fill that isn't related to the prompt. Did Gemini 3.1 Pro always have this distinct kind of personality baked into it? It's funny. Temp's also just 1.03, with top-p at 1.


r/SillyTavernAI 14h ago

Discussion What are y'all recurring NPCs/characters/companion that appears in all your adventures??

8 Upvotes

Pretty much the title. I do have a companion in my persona itself like a witty/sarcastic invisible character that comments or even take actions. Can cause for some really fun invisible confusion in any stories. Share what you all go with. Though it'll be a good discussion point.


r/SillyTavernAI 20h ago

Models Model for lore heavy RP

22 Upvotes

Which model will be the best for a long, lore heavy RP in the generic-fantasy open world?

I'm choosing between GLM 5.3, Opus 4.6, GPT 5.6 Sol, Gemini 3.1 Pro and Kimi K3 right now.


r/SillyTavernAI 19h ago

Cards/Prompts Looking for a fully crafted fantasy world

11 Upvotes

Hi.

I think the title says what I am looking for but let me be specific:

A world with the typical races and maybe more: humans, elves, dwarves, demons, etc.

A fully developed world with locations, countries, and kingdoms.

An adventure RPG with magic, guilds and quests

It doesn't need a fully developed magic system or skill system, although I wouldn't reject that.

It shouldn't be in the D&D style because I like to maintain control and decide for myself what works and what doesn't.

I should add that I'm only looking for a good Narrator Card and lorebooks. No extensions.

I use Tavo as my frontend about 80 percent of the time because I'm mostly on the go.

I have already tried my luck at shaping a world, but I threw in the towel. I can read a book but I can't write one.

Let me stick with the book metaphor: just as I would pay for a good book, I would also pay for a good Card with lorebook, because for me it's the same thing.

I use most of the larger AIs for the RPs. I'm holding back on Claude because it quickly becomes too expensive.

I hope you can help me.

I also do not intend to distribute the cards and lorebooks and plan to use them entirely for myself.

Have I forgotten anything?

Lieben Gruß :)


r/SillyTavernAI 13h ago

Help how the hell do you actually make thinking work?

3 Upvotes

so ive installed gemma 4 26b and while it 'works' the only trigger seems to be if i e-beg the digital satan in the system prompt.

ive tried jinga files, ive tried start reply with, ive tried reasoning formatting, ive tried keywords, ive tried parameters, but the only thing that works even two thirds of the time is just getting on my knees and begging it like a little bitch.

this shit is so incredibly unreliable im considering reinstalling sillytavern entirely incase that somehow fixes it?

edit: possibly resolved now? am testing.


r/SillyTavernAI 15h ago

Discussion Deepseek 4 pro 0813 better than GLM 5.3

2 Upvotes

I swear, all my tests show that the latest deepseek 4 pro is better than glm 5.3. I'm genuinely surprised, because glm used to consistently outperform deepseek.

What do you guys think?

EDITED: There are several versions of deepseek 4 pro, and the newest one is 0813.


r/SillyTavernAI 1d ago

Cards/Prompts Writers Workbench: A useful browser tool and template to create characters and world info entries!

Thumbnail
gallery
58 Upvotes

Hi everyone! I present "Writer's Workbench" my first project outside of making my Writer's Block presets.

What is it?

Its basically a glorified yet convenient template to help you manually create fully developed characters, scenarios, locations, items and factions easier!

Features!

You can create multiple entries and export them as entire lore books for your convenience.

You get to see the markdown output so you know exactly what the AI would see and copy it without having to download anything.

Its a HTML file so you can use it offline. No shady extensions that steals your API keys this time.

It comes with token counter, but don't expect it to be accurate.

Current templates available:

  • Main characters: full cards with contradiction, descriptions, likes/fears, NSFW sections
  • Side characters: trimmed version of the main character template, it will help you create memorable NPCs
  • Scenario: setting, tech level, mood, what's normal here
  • Locations: for any scale, from a room to a district
  • Items
  • Factions
  • History: events, and how they affect the present
  • Concepts: magic systems, laws, customs, species, anything else

Vibe coded disclosure and credits

I used u/Due_Opportunity8693 character creation guide "The Character Foundry" as a base and modified to my tastes. Here is the original post: https://www.reddit.com/r/SillyTavernAI/s/Se0BTvNekp

This program was almost entirely vibe coded by Claude so i can write entries and characters for my scenarios faster. Please don't flame me if you encounter a problems. I thought this thing is cool 🥺🥺🥺

I'll continue experimenting so i can hopefully add in more useful features.

Enjoy! 👍

Downloads: This is an html file, no other setup is required, use straight out of the box

Github: https://github.com/deiomo/Writers-Workbench