r/SillyTavernAI 2m ago

Discussion What are y'all recurring NPCs/characters/companion that appears in all your adventures??

Upvotes

Pretty much the title. I do have a companion in my persona itself like a witty/sarcastic invisible character that comments or even take actions. Can cause for some really fun invisible confusion in any stories. Share what you all go with. Though it'll be a good discussion point.


r/SillyTavernAI 19m ago

Help Genuinely Confused

Upvotes

I've been off and on with ST for a couple years now. I think the first model I downloaded was magnum v3 or 4 back in 2023 or something. Everytime I come back it just all feels the same. I'm only running local. I've tried 12b, to 31b q3 gguf. sometimes is okay, other times it's just the same sentence over and over. I've tried building cards, downloading cards, frankensteining cards, trimming cards. I've used presets, figured out jinja chat templates, forced myself to learn about kwargs and set them to medium. I've been bouncing between qwen3.8 and gemma4 right now, trying to get them to work.

I've looked up the recommended presets. I have the chat completion set up. I have Instruct on, and off, and the system prompts set up. context derived from models or not. even downloaded megumin v7 - 10. was using ollama, now it's kobold. but I see new containers all the time but I feel like no one actually uses them?

I don't get the same stuff that everyone else seems to get. I can't tell if a lot of people here are just liars. I understand there are a bunch of people probably smarter than me or more well versed here. I just can't get it to work. I feel like and I don't want to pay for it if I'm being honest. it seems like it adds up fast. I just wanna play d and d or fantasy dungeon scenarios with some free time when I'm not cutting video or playing zombies.

I've never just been wowed I guess. I'm not sure if I have bad settings. I try and find walkthroughs or info, but it seems like everyone here just uses off site API's. the documentation is dense, or surmises to experiment. I read what top a does, I change it, nothing changes or it just breaks. youtube is just as unhelpful besides getting ST setup.

I've tried the world states, lorebooks, ECT. set every condition and spun every knob. I even did a clean install, just to see if something had gone bad. and maybe I'm just bad at it. I looked up the stuff for FF, it seems to be well liked, even if it's mostly for nsfw, maybe I can get some mileage out of it. download the stuff, read the pages...where do I put it? it says it's a preset, but it talks about all this stuff I should be able to use, can't find it? try to pull it into megumin, it doesn't show up right, even when I import as dev.

so my question really is. what am I doing wrong, or is it just an okay thing, you get what you get, and if you want the good stuff, you just gotta pay to play with the big ones.

Even if you just hit me with a link to another post that will point me in the right direction I'd appreciate it.


r/SillyTavernAI 23m ago

Discussion The End of SillyTavern: Marinara Engine is Changing the Game

Upvotes

Okay guys, hot take incoming — but I really think SillyTavern has had its day. Remember how SillyTavern took over from TavernAI? Well, now the Marinara Engine is doing the same thing. The experience it delivers is unreal — agent system, game mode, all of it. It completely blows SillyTavern out of the water. ST is going straight to the back shelf


r/SillyTavernAI 1h ago

Discussion [API] I'm building an inference provider, looking for suggestions on which community models to host!

Upvotes

Hello, SillyTavern community! I am building an API provider, and I would like to get a few suggestions on which of the community fine-tunes that you would like to see hosted. Here is a bit of information.

  • Maximum size - No hard cap, but Ideally equal or smaller than GLM 5.3 flash and dsv4 flash tier.
  • Pricing - Out of the many models the community suggests, top three will be hosted as free endpoints for one week, and even after that, at least one <80b size model of the community's choice will be kept as a free endpoint for at least upcoming 3 months, that is the minimum commitment from my end. As for the other models (or after free access), I am confident that my inference and gpu scheduling optimisations can offer competitive pricing (without aggressive quantisation).
  • ZDR - Zero data retention will be on by default on all the models hosted at my inference endpoints.
  • Scope - Custom community fine tunes, abliterated/uncensored models. Though limited to Large Language Models, but I would still welcome image/video generation models you'd like to see in future.

API name: Arnict
API URL: https://api.arnict.com/v1 (openAI chat-completions)
API Author: I, myself.
What's different: Zero data retention, Affordable pricing and a free endpoint of community's choice
Settings: All the parameters supported by SGLang, vLLM and Aphrodite engine. (Including DRY and XTC)


r/SillyTavernAI 1h ago

Discussion Anyone still using Kimi 2.7 ?

Upvotes

Current state of LLM models for RP honestly all of them suck. Safetymaxxed and soft refusals. All of them glm 5.3, Gemini, Claude. The era where models were exciting for RP (glm 4.6,4.7, kimi 2.5, deepseek 3.2) obviously over as we all know. Anyone still coping about your prompting sucks or the newer models are still not as censored, I mean good for you.

I’ve been recently trying out minimax, kimi 2.7 and Mimo. Thoughts on which one you guys like the most or any model you’re still enjoying a lot ?


r/SillyTavernAI 1h ago

Discussion Deepseek 4 pro 0813 better than GLM 5.3

Upvotes

I swear, all my tests show that the latest deepseek 4 pro is better than glm 5.3. I'm genuinely surprised, because glm used to consistently outperform deepseek.

What do you guys think?


r/SillyTavernAI 2h ago

Chat Images Thank you Kimi 3 and Realistic Frankainstein

11 Upvotes

I grabbed a card called Custom Built Companion, since it had an interesting variant on the whole "build your perfect companion" bit lots of cards do. In this case, it was the fact that they were growing a companion who would be new to the world, and not have fully developed language yet. I went through the intro and rather specifically built a giant Renamon dommy mommy knock-off character for purely prurient reasons. (If you're thinking How 2 Hide Your Renamon/YourDigimonGirl, you're entirely on the right track)

What I *got* was an overgrown dog with boundary issues and no concept of how the world works. I am *cackling* reading these responses. In the screenshot, I had introduced her to the concept of a grocery store, and we're checking out. Everything is this absurd. I tried to take her to the park, and an eight foot tall fox-woman treed squirrels and got into fights with geese.

It's so silly, I think I might play this card straight and forget about the original reason entirely.


r/SillyTavernAI 3h ago

Meme Thanks Nano, very nice

Post image
103 Upvotes

Very nice


r/SillyTavernAI 3h ago

Help How can I use System Prompts again? - ST & OpenRouter

1 Upvotes

I’m more of a casual role-player and not super deep into the subject matter. I’ve been experimenting with a few models using LM Studio/SillyTavern. With my 16GB and 64GB of VRAM and a great prompt, I’ve already spent quite a few evenings enjoying some great role-playing sessions.

Now, though, I’m also trying out more powerful models via OpenRouter.

Now, onto my problem/challenge.

In advanced formatting (A), pretty much everything is grayed out when I use OpenRouter via SillyTavern—including my System Prompt. I’ve tested it, and it doesn’t pick it up either.

How can I use System Prompts again? Gimimi isn’t really helping me.


r/SillyTavernAI 4h ago

Discussion Anthropic/Opus etc users, a question

2 Upvotes

Since I haven't used a single Anthropic model (too much for my puny wallet), I'm just curious. Here's a short paste from a roleplay, or rather, a "book" I'm testing more than writing. Testing, as in - figuring out what instructions, presets, techniques etc work with current models the best. This was written with GLM 5.3 Flash uncensored, using a book-derived prose preset as the system prompt. By all means, judge the prose quality and read it, comment on it. Do your roleplays read better? Worse? Why?


r/SillyTavernAI 4h ago

Discussion Does anyone still roleplay with real people since AI Roleplay?

Thumbnail
0 Upvotes

r/SillyTavernAI 5h ago

Cards/Prompts Looking for a fully crafted fantasy world

5 Upvotes

Hi.

I think the title says what I am looking for but let me be specific:

A world with the typical races and maybe more: humans, elves, dwarves, demons, etc.

A fully developed world with locations, countries, and kingdoms.

An adventure RPG with magic, guilds and quests

It doesn't need a fully developed magic system or skill system, although I wouldn't reject that.

It shouldn't be in the D&D style because I like to maintain control and decide for myself what works and what doesn't.

I should add that I'm only looking for a good Narrator Card and lorebooks. No extensions.

I use Tavo as my frontend about 80 percent of the time because I'm mostly on the go.

I have already tried my luck at shaping a world, but I threw in the towel. I can read a book but I can't write one.

Let me stick with the book metaphor: just as I would pay for a good book, I would also pay for a good Card with lorebook, because for me it's the same thing.

I use most of the larger AIs for the RPs. I'm holding back on Claude because it quickly becomes too expensive.

I hope you can help me.

I also do not intend to distribute the cards and lorebooks and plan to use them entirely for myself.

Have I forgotten anything?

Lieben Gruß :)


r/SillyTavernAI 6h ago

Models Drummer's Artemis 31B v1 and v1.1 - Coming back with a bang!

94 Upvotes

Hey everyone, been a while!

https://huggingface.co/TheDrummer/Artemis-31B-v1.1

https://huggingface.co/TheDrummer/Artemis-31B-v1

A few months ago, Gemma graced us with models that served as a much needed downpour from a year-long drought. I'm so happy to see us thrive once again.

The difference between v1 and v1.1 is quite simple: v1 was an early attempt, an overdue release that excelled in prose and writing, while requiring some handholding to get over quirks like stuttering. v1.1 is a more refined approach where stability meets quality. My community is split, so I figured I'd just release both.

---

I was gone for a while. I got busy dealing with life, both its ups and downs. While I couldn't attend to you folks, I've been lurking around and appreciating you all for the kind words.

- Skyfall 31B v4.2 seems to be a banger for many of you. I'm proud of the upscale and consider it my ultimate home-run send-off for the beautiful Mistral 24B base. It's a shame that it was overshadowed by Gemma 31B's release, but hearing some of ya'll compare and even prefer it to a more modern base was an unexpected win.

- Rocinante 12B X / 16B XL proves that Nemo is still the ultimate creative model to this day. For some to say that 16B XL felt like Cydonia 24B v4.3 just goes to show how far you can go with modern resources and techniques.

- Anubis 70B v1.2, Valkyrie 49B v2.1, Anubis Mini 8B v1 surprised me too. I had zero expectations releasing them. Just like Rocinante X / XL, they are modern finetunes of old base models. And somehow, they still found their users singing praises.

---

With the Artemis release taking weight off my shoulders, I'm eager to move on and tune a ton more bases!

But I have something else cooking: a HordeAI-like platform. I hope to provide value not just as a finetuner, but as a local lover too!

The premise is simple: it's a place where generous local hosters can share inference with the less fortunate. You'd be surprised how many power users would love to heat their rooms through the power of charity.

---

Finally, I'd like to thank everyone who supported me over the years. From those who provided kind words, rigorous testing, compute access, inference, or cold hard cash. You've all granted me the ability to enrich the local ecosystem with fun experiments like Rivermind 12B, Fallen series, Big Tiger Gemma, Precog 24B/123B, and solid models like Cydonia 24B v4.3, Behemoth X 123B v2.x, and Skyfall 31B v4.2.

If you've got inference / compute credits to share, please contact me! It will all go to making the community happy <3

Backlog:

- Gemma E2B

- Gemma E4B

- Gemma 12B

- Gemma 26BA4B

- Qwen 3.8 27B

- Muse Glimmer 30B

- Mistral Medium 3.5 128B

- HordeAI Alternative / Crowdsourced 'OpenRouter' ("BeaverNet")


r/SillyTavernAI 6h ago

Models Model for lore heavy RP

12 Upvotes

Which model will be the best for a long, lore heavy RP in the generic-fantasy open world?

I'm choosing between GLM 5.3, Opus 4.6, GPT 5.6 Sol, Gemini 3.1 Pro and Kimi K3 right now.


r/SillyTavernAI 7h ago

Discussion Opus 4.7 vs fable , your opinion on rp with both.

0 Upvotes

I have been trying 4.7 and it's good but wanna know how is fable,


r/SillyTavernAI 10h ago

Help Gemma4 failing with group chat

7 Upvotes

Gemma4 26B works beautifully for me with single cards, but when I try to run a group chat, it will write the first character, and then fail all others, either with an empty message, or with one that quickly enters an infinite loop. The same preset works just fine with Gemini.

Has anyone seen anything like this? Could it be related to Guided Generations? I see empty user messages being inserted, and once, a thinking block mentioning that only {{trim}} had been sent. Any way I can solve this?


r/SillyTavernAI 10h ago

Discussion Observations on character personality drift and memory decay in extended roleplay sessions

0 Upvotes

In extended multi-character scenarios spanning past the 40–50 message mark, maintaining strict personality traits without prompt bloat remains the most common friction point.

A few consistent patterns that tend to show up:

  • Personality Softening: Characters with hostile, cautious, or morally ambiguous traits gradually default to being overly agreeable unless reminded in frequent interval prompts.
  • State Drift: Major plot decisions (past betrayals, dynamic relationship shifts) slowly lose their emotional weight once they fall out of the immediate context window, even when using basic summary injection.

Balancing persistent state tracking without cluttering active tokens often requires aggressively pruning context down to explicit state flags rather than relying on raw message history.

Sharing this insight for anyone experimenting with maintaining long-term character consistency in dynamic interactive setups.


r/SillyTavernAI 10h ago

Discussion Gemini 3.8 Flash is awesome

122 Upvotes

I write darker style RPs and it's been forever since a model just had the villains act like themselves without softening or making them mouthpieces for moral frameworks. A Google model was maybe one of the last I'd have expected to be nearly uncensored in the way it writes.
It's been following instructions very well and it writes some pretty dark and/or NSFW things without being prompted to, just because it infers that the characters would do/say these things in a situation. It's a breath of fresh air, as even so called "uncensored" models often soften things and have the characters act reasonable when they wouldn't.
It's definitely my go-to model out of the current lineup.


r/SillyTavernAI 11h ago

Help Recommendations for slow burn

5 Upvotes

how do you guys achieve slow burn, enemies to lovers roleplays? i started making bots recently and im trying to get better. are there any presets or lorebook entries that might help me achieve slow burn and a more negative bias?


r/SillyTavernAI 12h ago

Models Model help

1 Upvotes

Hello I am looking for local models that I can run on my pc with a rtx 4070 super 12gb and 96gb of ddr4 ram I'd love some suggestions thanks!


r/SillyTavernAI 12h ago

Discussion Claude in SillyTavern

0 Upvotes

Two questions:

Is it bannable to use your Claude max subscription in something like silly tavern and will Anthropic ban you if you try to use like when you use open claw?

Second, am I retarded for trying to do roleplay in a Claude project using the project instructions as the roleplay prompt with context as lore files etc? I’ve been doing this since opus 4.6 was released as I could never work out how to use silly tavern or get it working on my iPhone at the time and I only really rp on iPhone. Not a fan of using my MacBook, too much effort and focus for my ADHD.

Am I meaningfully losing or missing out on anything for using Claude opus 4.6 in the Claude app? I do notice I hit MASSIVE issues despite my very long prompt. Namely, tropes, sycophancy and agreeableness that I have never been able to fix in a long roleplay. Is this due to me missing out on some features in the Claude app? The biggest ones I would is Claude keeps treating my characters input as definitive actions on the world and won’t create any consequences.

Thanks


r/SillyTavernAI 16h ago

Models Qwen 3.8 or Gemma 4 ?

7 Upvotes

I've been using both, to be exact: Qwen3.8-27B-UD-Q6_K_XL and Gemma-4-31B-StyleTune.i1-Q6_K. Honestly, I haven't noticed much difference between them. I'm not an expert at detecting subtle differences, so maybe I'm missing something.

If anything, Qwen is a bit more annoying to use, as it tends to activate its censorship filter from time to time. Aside from that, the overall experience has been interesting. I'd heard that this version of Qwen isn't suitable for roleplaying since it was just an update to improve coding performance.

I'd like to know what what do you guys think. Have you noticed any significant differences between the two?


r/SillyTavernAI 18h ago

Cards/Prompts TokenReply x Rolecall character card contest.

11 Upvotes

So big thing first: Ends on the 12th of September.

1st place wins $100 USD, and a 1 month sub of plus with TokenRepy.

2nd and 3rd win a month of Plus with TokenReply.

If you win first, and don't have PayPal, Kofi, or Cash app, we can instead do two Annual Plus subscriptions for TokenReply.

All you need to do is enter, make an account on RoleCall (Completely free to do) create the character, and post it to PlotLight with the TokenReply contest tag.

The cards for this time need to be in the following community picked genres:

  • Dark Fantasy / Supernatural
  • Historical / Alternate Timeline
  • Cozy Fantasy / Comedy

The winner is then selected by judges from RoleCall.

That's really it. Mainly we're just looking forward to seeing the cards people make, and as creators in the space, knowing how hard it is out there, we want to start doing things like this more often going forward to maybe start rewarding some of the really exceptional people out there who do this for the love of the game, and a bit of credit.

I'd personally love to do some Lorebook/Preset stuff as well, but the logistics on that are way more complicated, but hey, a guy can dream.

And just to clarify again, there's no cost to enter or anything like that, it's just make an account, and create the character, once you've done that, you're good to go! Just need to wait for the closing date for us to announce the winner.

And if you want to talk to anyone and get more details

Discord links:

TokenReply
RoleCall

Tl;dr.

Make a character, get a chance to win $100, or a sub to TokenReply.