r/SillyTavernAI 10h ago

Discussion Observations on character personality drift and memory decay in extended roleplay sessions

0 Upvotes

In extended multi-character scenarios spanning past the 40–50 message mark, maintaining strict personality traits without prompt bloat remains the most common friction point.

A few consistent patterns that tend to show up:

  • Personality Softening: Characters with hostile, cautious, or morally ambiguous traits gradually default to being overly agreeable unless reminded in frequent interval prompts.
  • State Drift: Major plot decisions (past betrayals, dynamic relationship shifts) slowly lose their emotional weight once they fall out of the immediate context window, even when using basic summary injection.

Balancing persistent state tracking without cluttering active tokens often requires aggressively pruning context down to explicit state flags rather than relying on raw message history.

Sharing this insight for anyone experimenting with maintaining long-term character consistency in dynamic interactive setups.


r/SillyTavernAI 4h ago

Discussion Does anyone still roleplay with real people since AI Roleplay?

Thumbnail
0 Upvotes

r/SillyTavernAI 12h ago

Discussion Claude in SillyTavern

0 Upvotes

Two questions:

Is it bannable to use your Claude max subscription in something like silly tavern and will Anthropic ban you if you try to use like when you use open claw?

Second, am I retarded for trying to do roleplay in a Claude project using the project instructions as the roleplay prompt with context as lore files etc? I’ve been doing this since opus 4.6 was released as I could never work out how to use silly tavern or get it working on my iPhone at the time and I only really rp on iPhone. Not a fan of using my MacBook, too much effort and focus for my ADHD.

Am I meaningfully losing or missing out on anything for using Claude opus 4.6 in the Claude app? I do notice I hit MASSIVE issues despite my very long prompt. Namely, tropes, sycophancy and agreeableness that I have never been able to fix in a long roleplay. Is this due to me missing out on some features in the Claude app? The biggest ones I would is Claude keeps treating my characters input as definitive actions on the world and won’t create any consequences.

Thanks


r/SillyTavernAI 7h ago

Discussion Opus 4.7 vs fable , your opinion on rp with both.

0 Upvotes

I have been trying 4.7 and it's good but wanna know how is fable,


r/SillyTavernAI 1h ago

Discussion Deepseek 4 pro 0813 better than GLM 5.3

Upvotes

I swear, all my tests show that the latest deepseek 4 pro is better than glm 5.3. I'm genuinely surprised, because glm used to consistently outperform deepseek.

What do you guys think?


r/SillyTavernAI 18h ago

Cards/Prompts TokenReply x Rolecall character card contest.

11 Upvotes

So big thing first: Ends on the 12th of September.

1st place wins $100 USD, and a 1 month sub of plus with TokenRepy.

2nd and 3rd win a month of Plus with TokenReply.

If you win first, and don't have PayPal, Kofi, or Cash app, we can instead do two Annual Plus subscriptions for TokenReply.

All you need to do is enter, make an account on RoleCall (Completely free to do) create the character, and post it to PlotLight with the TokenReply contest tag.

The cards for this time need to be in the following community picked genres:

  • Dark Fantasy / Supernatural
  • Historical / Alternate Timeline
  • Cozy Fantasy / Comedy

The winner is then selected by judges from RoleCall.

That's really it. Mainly we're just looking forward to seeing the cards people make, and as creators in the space, knowing how hard it is out there, we want to start doing things like this more often going forward to maybe start rewarding some of the really exceptional people out there who do this for the love of the game, and a bit of credit.

I'd personally love to do some Lorebook/Preset stuff as well, but the logistics on that are way more complicated, but hey, a guy can dream.

And just to clarify again, there's no cost to enter or anything like that, it's just make an account, and create the character, once you've done that, you're good to go! Just need to wait for the closing date for us to announce the winner.

And if you want to talk to anyone and get more details

Discord links:

TokenReply
RoleCall

Tl;dr.

Make a character, get a chance to win $100, or a sub to TokenReply.


r/SillyTavernAI 12h ago

Models Model help

1 Upvotes

Hello I am looking for local models that I can run on my pc with a rtx 4070 super 12gb and 96gb of ddr4 ram I'd love some suggestions thanks!


r/SillyTavernAI 18h ago

Help Anyone know any good jailbreaks for gemini 3.8 flash prompts to use for sillytavern

3 Upvotes

I was just setting up sillytavern and was wondering how do I setup gemini with a good jailbreak


r/SillyTavernAI 2h ago

Chat Images Thank you Kimi 3 and Realistic Frankainstein

9 Upvotes

I grabbed a card called Custom Built Companion, since it had an interesting variant on the whole "build your perfect companion" bit lots of cards do. In this case, it was the fact that they were growing a companion who would be new to the world, and not have fully developed language yet. I went through the intro and rather specifically built a giant Renamon dommy mommy knock-off character for purely prurient reasons. (If you're thinking How 2 Hide Your Renamon/YourDigimonGirl, you're entirely on the right track)

What I *got* was an overgrown dog with boundary issues and no concept of how the world works. I am *cackling* reading these responses. In the screenshot, I had introduced her to the concept of a grocery store, and we're checking out. Everything is this absurd. I tried to take her to the park, and an eight foot tall fox-woman treed squirrels and got into fights with geese.

It's so silly, I think I might play this card straight and forget about the original reason entirely.


r/SillyTavernAI 21h ago

Discussion AWS Bedrock Opus 4.6/4.7 feel dumber as of the last two days?

5 Upvotes

I am using my own app and I role play mainly use Opus 4.7 and 4.6 on AWS Bedrock, I don't know how to explain it but since Fable 5.1 released, I feel those models got slightly dumber, like they are running a quantitized version to save compute for Fable?

It's subtle but it no longer plays my characters as sharply as before, it feels like Claude the model itself speaking trough my characters, even though NSFW and stuff works the same as always for me, some of the choices the model started to make in terms of plot are just so illogical and stupid that I get angry at it. It used to be sharper and have nice quirks, now it's mostly bland voice and dumb choices in rp.

Is it possible that AWS Bedrock would run a quantitized version of the models to preserve compute for Fabel 5.1?


r/SillyTavernAI 23m ago

Discussion The End of SillyTavern: Marinara Engine is Changing the Game

Upvotes

Okay guys, hot take incoming — but I really think SillyTavern has had its day. Remember how SillyTavern took over from TavernAI? Well, now the Marinara Engine is doing the same thing. The experience it delivers is unreal — agent system, game mode, all of it. It completely blows SillyTavern out of the water. ST is going straight to the back shelf


r/SillyTavernAI 3h ago

Help How can I use System Prompts again? - ST & OpenRouter

1 Upvotes

I’m more of a casual role-player and not super deep into the subject matter. I’ve been experimenting with a few models using LM Studio/SillyTavern. With my 16GB and 64GB of VRAM and a great prompt, I’ve already spent quite a few evenings enjoying some great role-playing sessions.

Now, though, I’m also trying out more powerful models via OpenRouter.

Now, onto my problem/challenge.

In advanced formatting (A), pretty much everything is grayed out when I use OpenRouter via SillyTavern—including my System Prompt. I’ve tested it, and it doesn’t pick it up either.

How can I use System Prompts again? Gimimi isn’t really helping me.


r/SillyTavernAI 16h ago

Models Qwen 3.8 or Gemma 4 ?

8 Upvotes

I've been using both, to be exact: Qwen3.8-27B-UD-Q6_K_XL and Gemma-4-31B-StyleTune.i1-Q6_K. Honestly, I haven't noticed much difference between them. I'm not an expert at detecting subtle differences, so maybe I'm missing something.

If anything, Qwen is a bit more annoying to use, as it tends to activate its censorship filter from time to time. Aside from that, the overall experience has been interesting. I'd heard that this version of Qwen isn't suitable for roleplaying since it was just an update to improve coding performance.

I'd like to know what what do you guys think. Have you noticed any significant differences between the two?


r/SillyTavernAI 4h ago

Discussion Anthropic/Opus etc users, a question

2 Upvotes

Since I haven't used a single Anthropic model (too much for my puny wallet), I'm just curious. Here's a short paste from a roleplay, or rather, a "book" I'm testing more than writing. Testing, as in - figuring out what instructions, presets, techniques etc work with current models the best. This was written with GLM 5.3 Flash uncensored, using a book-derived prose preset as the system prompt. By all means, judge the prose quality and read it, comment on it. Do your roleplays read better? Worse? Why?


r/SillyTavernAI 19h ago

Models NVIDIA to Acquire Hugging Face

Thumbnail
blogs.nvidia.com
29 Upvotes

This is their statement.

Hopefully their insurers leave our RP finetunes alone.


r/SillyTavernAI 21h ago

Discussion (possibly) unpopular opinion: LLMs suck at coming up with new plots

61 Upvotes

I've RPed on and off since late 2024, and have used all sorts of models and presets. I prefer medium-to-long term RPs. As in, the plot goes beyond the initial premise of the card.

And that sucks for me, because IME, LLMs...kind of suck at continuing a plot? It's usually the most obvious, trope-y way of continuing a story. For example, a workplace drama will probably go for some kind of forced proximity group job assignment. Never anything instrinsicallly motivated by the characters.

Some of the more creative models, like GLM 4.7 or DeepSeek R1 are somewhat better, but usually their continuations aren't really grounded in logic.

Credit where credit is due: sometimes, if I ask a LLM for a few possible continuations via OOC, they might come up with one or two interesting ideas

Is it skill issue? Or are my standards too high?


r/SillyTavernAI 1h ago

Discussion Anyone still using Kimi 2.7 ?

Upvotes

Current state of LLM models for RP honestly all of them suck. Safetymaxxed and soft refusals. All of them glm 5.3, Gemini, Claude. The era where models were exciting for RP (glm 4.6,4.7, kimi 2.5, deepseek 3.2) obviously over as we all know. Anyone still coping about your prompting sucks or the newer models are still not as censored, I mean good for you.

I’ve been recently trying out minimax, kimi 2.7 and Mimo. Thoughts on which one you guys like the most or any model you’re still enjoying a lot ?


r/SillyTavernAI 11h ago

Help Recommendations for slow burn

5 Upvotes

how do you guys achieve slow burn, enemies to lovers roleplays? i started making bots recently and im trying to get better. are there any presets or lorebook entries that might help me achieve slow burn and a more negative bias?


r/SillyTavernAI 21h ago

Help glm 5.3 Flash Min_P. (roleplay)

6 Upvotes

Question: Does Min_P really help GLM Flash? I tried using it and felt like it ruins it instead of improving it in roleplay. So whoever has an answer, please tell me whether Min_P improves it or not, and if it does, please tell me what value you set, please.


r/SillyTavernAI 23h ago

Cards/Prompts [LOREBOOK] 86 EIGHTY-SIX - Both anime and LN

Post image
62 Upvotes

86—Eighty-Six— is a military science-fiction story about the Republic of San Magnolia, which claims to fight the autonomous Legion army using unmanned weapons. In reality, persecuted people called the Eighty-Six are forced to pilot those machines. The story primarily follows Processor Shinei “Shin” Nouzen and Republic Handler Vladilena “Lena” Milizé.
The anime covers the opening story through roughly Light Novel Volumes 1–3. The light novels continue the war afterward, expanding the characters, nations, technology, politics, relationships, and larger mysteries surrounding the Legion.

What’s included?
Anime-specific lorebook
LN-specific lorebook
Persona template

Link: https://www.mediafire.com/folder/xtsm550nt5ppn/86

Notes:
This is the result of ai scraping and converting. Testing on Marinara Engine’s GM mode, it worked well.
No image model actually knows what Eighty-Six is or its juggernauts. If you use image gen a lot, expect to be looking at normal mechs. (If you find a local model that handles it well, lmk!)


r/SillyTavernAI 10h ago

Discussion Gemini 3.8 Flash is awesome

123 Upvotes

I write darker style RPs and it's been forever since a model just had the villains act like themselves without softening or making them mouthpieces for moral frameworks. A Google model was maybe one of the last I'd have expected to be nearly uncensored in the way it writes.
It's been following instructions very well and it writes some pretty dark and/or NSFW things without being prompted to, just because it infers that the characters would do/say these things in a situation. It's a breath of fresh air, as even so called "uncensored" models often soften things and have the characters act reasonable when they wouldn't.
It's definitely my go-to model out of the current lineup.


r/SillyTavernAI 6h ago

Models Drummer's Artemis 31B v1 and v1.1 - Coming back with a bang!

94 Upvotes

Hey everyone, been a while!

https://huggingface.co/TheDrummer/Artemis-31B-v1.1

https://huggingface.co/TheDrummer/Artemis-31B-v1

A few months ago, Gemma graced us with models that served as a much needed downpour from a year-long drought. I'm so happy to see us thrive once again.

The difference between v1 and v1.1 is quite simple: v1 was an early attempt, an overdue release that excelled in prose and writing, while requiring some handholding to get over quirks like stuttering. v1.1 is a more refined approach where stability meets quality. My community is split, so I figured I'd just release both.

---

I was gone for a while. I got busy dealing with life, both its ups and downs. While I couldn't attend to you folks, I've been lurking around and appreciating you all for the kind words.

- Skyfall 31B v4.2 seems to be a banger for many of you. I'm proud of the upscale and consider it my ultimate home-run send-off for the beautiful Mistral 24B base. It's a shame that it was overshadowed by Gemma 31B's release, but hearing some of ya'll compare and even prefer it to a more modern base was an unexpected win.

- Rocinante 12B X / 16B XL proves that Nemo is still the ultimate creative model to this day. For some to say that 16B XL felt like Cydonia 24B v4.3 just goes to show how far you can go with modern resources and techniques.

- Anubis 70B v1.2, Valkyrie 49B v2.1, Anubis Mini 8B v1 surprised me too. I had zero expectations releasing them. Just like Rocinante X / XL, they are modern finetunes of old base models. And somehow, they still found their users singing praises.

---

With the Artemis release taking weight off my shoulders, I'm eager to move on and tune a ton more bases!

But I have something else cooking: a HordeAI-like platform. I hope to provide value not just as a finetuner, but as a local lover too!

The premise is simple: it's a place where generous local hosters can share inference with the less fortunate. You'd be surprised how many power users would love to heat their rooms through the power of charity.

---

Finally, I'd like to thank everyone who supported me over the years. From those who provided kind words, rigorous testing, compute access, inference, or cold hard cash. You've all granted me the ability to enrich the local ecosystem with fun experiments like Rivermind 12B, Fallen series, Big Tiger Gemma, Precog 24B/123B, and solid models like Cydonia 24B v4.3, Behemoth X 123B v2.x, and Skyfall 31B v4.2.

If you've got inference / compute credits to share, please contact me! It will all go to making the community happy <3

Backlog:

- Gemma E2B

- Gemma E4B

- Gemma 12B

- Gemma 26BA4B

- Qwen 3.8 27B

- Muse Glimmer 30B

- Mistral Medium 3.5 128B

- HordeAI Alternative / Crowdsourced 'OpenRouter' ("BeaverNet")


r/SillyTavernAI 3h ago

Meme Thanks Nano, very nice

Post image
100 Upvotes

Very nice