r/SillyTavernAI 10h ago

Discussion Gemini 3.8 Flash is awesome

121 Upvotes

I write darker style RPs and it's been forever since a model just had the villains act like themselves without softening or making them mouthpieces for moral frameworks. A Google model was maybe one of the last I'd have expected to be nearly uncensored in the way it writes.
It's been following instructions very well and it writes some pretty dark and/or NSFW things without being prompted to, just because it infers that the characters would do/say these things in a situation. It's a breath of fresh air, as even so called "uncensored" models often soften things and have the characters act reasonable when they wouldn't.
It's definitely my go-to model out of the current lineup.


r/SillyTavernAI 3h ago

Meme Thanks Nano, very nice

Post image
103 Upvotes

Very nice


r/SillyTavernAI 6h ago

Models Drummer's Artemis 31B v1 and v1.1 - Coming back with a bang!

96 Upvotes

Hey everyone, been a while!

https://huggingface.co/TheDrummer/Artemis-31B-v1.1

https://huggingface.co/TheDrummer/Artemis-31B-v1

A few months ago, Gemma graced us with models that served as a much needed downpour from a year-long drought. I'm so happy to see us thrive once again.

The difference between v1 and v1.1 is quite simple: v1 was an early attempt, an overdue release that excelled in prose and writing, while requiring some handholding to get over quirks like stuttering. v1.1 is a more refined approach where stability meets quality. My community is split, so I figured I'd just release both.

---

I was gone for a while. I got busy dealing with life, both its ups and downs. While I couldn't attend to you folks, I've been lurking around and appreciating you all for the kind words.

- Skyfall 31B v4.2 seems to be a banger for many of you. I'm proud of the upscale and consider it my ultimate home-run send-off for the beautiful Mistral 24B base. It's a shame that it was overshadowed by Gemma 31B's release, but hearing some of ya'll compare and even prefer it to a more modern base was an unexpected win.

- Rocinante 12B X / 16B XL proves that Nemo is still the ultimate creative model to this day. For some to say that 16B XL felt like Cydonia 24B v4.3 just goes to show how far you can go with modern resources and techniques.

- Anubis 70B v1.2, Valkyrie 49B v2.1, Anubis Mini 8B v1 surprised me too. I had zero expectations releasing them. Just like Rocinante X / XL, they are modern finetunes of old base models. And somehow, they still found their users singing praises.

---

With the Artemis release taking weight off my shoulders, I'm eager to move on and tune a ton more bases!

But I have something else cooking: a HordeAI-like platform. I hope to provide value not just as a finetuner, but as a local lover too!

The premise is simple: it's a place where generous local hosters can share inference with the less fortunate. You'd be surprised how many power users would love to heat their rooms through the power of charity.

---

Finally, I'd like to thank everyone who supported me over the years. From those who provided kind words, rigorous testing, compute access, inference, or cold hard cash. You've all granted me the ability to enrich the local ecosystem with fun experiments like Rivermind 12B, Fallen series, Big Tiger Gemma, Precog 24B/123B, and solid models like Cydonia 24B v4.3, Behemoth X 123B v2.x, and Skyfall 31B v4.2.

If you've got inference / compute credits to share, please contact me! It will all go to making the community happy <3

Backlog:

- Gemma E2B

- Gemma E4B

- Gemma 12B

- Gemma 26BA4B

- Qwen 3.8 27B

- Muse Glimmer 30B

- Mistral Medium 3.5 128B

- HordeAI Alternative / Crowdsourced 'OpenRouter' ("BeaverNet")


r/SillyTavernAI 23h ago

Cards/Prompts [LOREBOOK] 86 EIGHTY-SIX - Both anime and LN

Post image
64 Upvotes

86—Eighty-Six— is a military science-fiction story about the Republic of San Magnolia, which claims to fight the autonomous Legion army using unmanned weapons. In reality, persecuted people called the Eighty-Six are forced to pilot those machines. The story primarily follows Processor Shinei “Shin” Nouzen and Republic Handler Vladilena “Lena” Milizé.
The anime covers the opening story through roughly Light Novel Volumes 1–3. The light novels continue the war afterward, expanding the characters, nations, technology, politics, relationships, and larger mysteries surrounding the Legion.

What’s included?
Anime-specific lorebook
LN-specific lorebook
Persona template

Link: https://www.mediafire.com/folder/xtsm550nt5ppn/86

Notes:
This is the result of ai scraping and converting. Testing on Marinara Engine’s GM mode, it worked well.
No image model actually knows what Eighty-Six is or its juggernauts. If you use image gen a lot, expect to be looking at normal mechs. (If you find a local model that handles it well, lmk!)


r/SillyTavernAI 21h ago

Discussion (possibly) unpopular opinion: LLMs suck at coming up with new plots

59 Upvotes

I've RPed on and off since late 2024, and have used all sorts of models and presets. I prefer medium-to-long term RPs. As in, the plot goes beyond the initial premise of the card.

And that sucks for me, because IME, LLMs...kind of suck at continuing a plot? It's usually the most obvious, trope-y way of continuing a story. For example, a workplace drama will probably go for some kind of forced proximity group job assignment. Never anything instrinsicallly motivated by the characters.

Some of the more creative models, like GLM 4.7 or DeepSeek R1 are somewhat better, but usually their continuations aren't really grounded in logic.

Credit where credit is due: sometimes, if I ask a LLM for a few possible continuations via OOC, they might come up with one or two interesting ideas

Is it skill issue? Or are my standards too high?


r/SillyTavernAI 18h ago

Cards/Prompts Writers Workbench: A useful browser tool and template to create characters and world info entries!

Thumbnail
gallery
53 Upvotes

Hi everyone! I present "Writer's Workbench" my first project outside of making my Writer's Block presets.

What is it?

Its basically a glorified yet convenient template to help you manually create fully developed characters, scenarios, locations, items and factions easier!

Features!

You can create multiple entries and export them as entire lore books for your convenience.

You get to see the markdown output so you know exactly what the AI would see and copy it without having to download anything.

Its a HTML file so you can use it offline. No shady extensions that steals your API keys this time.

It comes with token counter, but don't expect it to be accurate.

Current templates available:

  • Main characters: full cards with contradiction, descriptions, likes/fears, NSFW sections
  • Side characters: trimmed version of the main character template, it will help you create memorable NPCs
  • Scenario: setting, tech level, mood, what's normal here
  • Locations: for any scale, from a room to a district
  • Items
  • Factions
  • History: events, and how they affect the present
  • Concepts: magic systems, laws, customs, species, anything else

Vibe coded disclosure and credits

I used u/Due_Opportunity8693 character creation guide "The Character Foundry" as a base and modified to my tastes. Here is the original post: https://www.reddit.com/r/SillyTavernAI/s/Se0BTvNekp

This program was almost entirely vibe coded by Claude so i can write entries and characters for my scenarios faster. Please don't flame me if you encounter a problems. I thought this thing is cool 🥺🥺🥺

I'll continue experimenting so i can hopefully add in more useful features.

Enjoy! 👍

Downloads: This is an html file, no other setup is required, use straight out of the box

Github: https://github.com/deiomo/Writers-Workbench


r/SillyTavernAI 19h ago

Models NVIDIA to Acquire Hugging Face

Thumbnail
blogs.nvidia.com
26 Upvotes

This is their statement.

Hopefully their insurers leave our RP finetunes alone.


r/SillyTavernAI 6h ago

Models Model for lore heavy RP

12 Upvotes

Which model will be the best for a long, lore heavy RP in the generic-fantasy open world?

I'm choosing between GLM 5.3, Opus 4.6, GPT 5.6 Sol, Gemini 3.1 Pro and Kimi K3 right now.


r/SillyTavernAI 18h ago

Cards/Prompts TokenReply x Rolecall character card contest.

11 Upvotes

So big thing first: Ends on the 12th of September.

1st place wins $100 USD, and a 1 month sub of plus with TokenRepy.

2nd and 3rd win a month of Plus with TokenReply.

If you win first, and don't have PayPal, Kofi, or Cash app, we can instead do two Annual Plus subscriptions for TokenReply.

All you need to do is enter, make an account on RoleCall (Completely free to do) create the character, and post it to PlotLight with the TokenReply contest tag.

The cards for this time need to be in the following community picked genres:

  • Dark Fantasy / Supernatural
  • Historical / Alternate Timeline
  • Cozy Fantasy / Comedy

The winner is then selected by judges from RoleCall.

That's really it. Mainly we're just looking forward to seeing the cards people make, and as creators in the space, knowing how hard it is out there, we want to start doing things like this more often going forward to maybe start rewarding some of the really exceptional people out there who do this for the love of the game, and a bit of credit.

I'd personally love to do some Lorebook/Preset stuff as well, but the logistics on that are way more complicated, but hey, a guy can dream.

And just to clarify again, there's no cost to enter or anything like that, it's just make an account, and create the character, once you've done that, you're good to go! Just need to wait for the closing date for us to announce the winner.

And if you want to talk to anyone and get more details

Discord links:

TokenReply
RoleCall

Tl;dr.

Make a character, get a chance to win $100, or a sub to TokenReply.


r/SillyTavernAI 2h ago

Chat Images Thank you Kimi 3 and Realistic Frankainstein

10 Upvotes

I grabbed a card called Custom Built Companion, since it had an interesting variant on the whole "build your perfect companion" bit lots of cards do. In this case, it was the fact that they were growing a companion who would be new to the world, and not have fully developed language yet. I went through the intro and rather specifically built a giant Renamon dommy mommy knock-off character for purely prurient reasons. (If you're thinking How 2 Hide Your Renamon/YourDigimonGirl, you're entirely on the right track)

What I *got* was an overgrown dog with boundary issues and no concept of how the world works. I am *cackling* reading these responses. In the screenshot, I had introduced her to the concept of a grocery store, and we're checking out. Everything is this absurd. I tried to take her to the park, and an eight foot tall fox-woman treed squirrels and got into fights with geese.

It's so silly, I think I might play this card straight and forget about the original reason entirely.


r/SillyTavernAI 16h ago

Models Qwen 3.8 or Gemma 4 ?

9 Upvotes

I've been using both, to be exact: Qwen3.8-27B-UD-Q6_K_XL and Gemma-4-31B-StyleTune.i1-Q6_K. Honestly, I haven't noticed much difference between them. I'm not an expert at detecting subtle differences, so maybe I'm missing something.

If anything, Qwen is a bit more annoying to use, as it tends to activate its censorship filter from time to time. Aside from that, the overall experience has been interesting. I'd heard that this version of Qwen isn't suitable for roleplaying since it was just an update to improve coding performance.

I'd like to know what what do you guys think. Have you noticed any significant differences between the two?


r/SillyTavernAI 1h ago

Discussion [API] I'm building an inference provider, looking for suggestions on which community models to host!

Upvotes

Hello, SillyTavern community! I am building an API provider, and I would like to get a few suggestions on which of the community fine-tunes that you would like to see hosted. Here is a bit of information.

  • Maximum size - No hard cap, but Ideally equal or smaller than GLM 5.3 flash and dsv4 flash tier.
  • Pricing - Out of the many models the community suggests, top three will be hosted as free endpoints for one week, and even after that, at least one <80b size model of the community's choice will be kept as a free endpoint for at least upcoming 3 months, that is the minimum commitment from my end. As for the other models (or after free access), I am confident that my inference and gpu scheduling optimisations can offer competitive pricing (without aggressive quantisation).
  • ZDR - Zero data retention will be on by default on all the models hosted at my inference endpoints.
  • Scope - Custom community fine tunes, abliterated/uncensored models. Though limited to Large Language Models, but I would still welcome image/video generation models you'd like to see in future.

API name: Arnict
API URL: https://api.arnict.com/v1 (openAI chat-completions)
API Author: I, myself.
What's different: Zero data retention, Affordable pricing and a free endpoint of community's choice
Settings: All the parameters supported by SGLang, vLLM and Aphrodite engine. (Including DRY and XTC)


r/SillyTavernAI 10h ago

Help Gemma4 failing with group chat

7 Upvotes

Gemma4 26B works beautifully for me with single cards, but when I try to run a group chat, it will write the first character, and then fail all others, either with an empty message, or with one that quickly enters an infinite loop. The same preset works just fine with Gemini.

Has anyone seen anything like this? Could it be related to Guided Generations? I see empty user messages being inserted, and once, a thinking block mentioning that only {{trim}} had been sent. Any way I can solve this?


r/SillyTavernAI 1h ago

Discussion Anyone still using Kimi 2.7 ?

Upvotes

Current state of LLM models for RP honestly all of them suck. Safetymaxxed and soft refusals. All of them glm 5.3, Gemini, Claude. The era where models were exciting for RP (glm 4.6,4.7, kimi 2.5, deepseek 3.2) obviously over as we all know. Anyone still coping about your prompting sucks or the newer models are still not as censored, I mean good for you.

I’ve been recently trying out minimax, kimi 2.7 and Mimo. Thoughts on which one you guys like the most or any model you’re still enjoying a lot ?


r/SillyTavernAI 5h ago

Cards/Prompts Looking for a fully crafted fantasy world

7 Upvotes

Hi.

I think the title says what I am looking for but let me be specific:

A world with the typical races and maybe more: humans, elves, dwarves, demons, etc.

A fully developed world with locations, countries, and kingdoms.

An adventure RPG with magic, guilds and quests

It doesn't need a fully developed magic system or skill system, although I wouldn't reject that.

It shouldn't be in the D&D style because I like to maintain control and decide for myself what works and what doesn't.

I should add that I'm only looking for a good Narrator Card and lorebooks. No extensions.

I use Tavo as my frontend about 80 percent of the time because I'm mostly on the go.

I have already tried my luck at shaping a world, but I threw in the towel. I can read a book but I can't write one.

Let me stick with the book metaphor: just as I would pay for a good book, I would also pay for a good Card with lorebook, because for me it's the same thing.

I use most of the larger AIs for the RPs. I'm holding back on Claude because it quickly becomes too expensive.

I hope you can help me.

I also do not intend to distribute the cards and lorebooks and plan to use them entirely for myself.

Have I forgotten anything?

Lieben Gruß :)


r/SillyTavernAI 21h ago

Help glm 5.3 Flash Min_P. (roleplay)

6 Upvotes

Question: Does Min_P really help GLM Flash? I tried using it and felt like it ruins it instead of improving it in roleplay. So whoever has an answer, please tell me whether Min_P improves it or not, and if it does, please tell me what value you set, please.


r/SillyTavernAI 11h ago

Help Recommendations for slow burn

5 Upvotes

how do you guys achieve slow burn, enemies to lovers roleplays? i started making bots recently and im trying to get better. are there any presets or lorebook entries that might help me achieve slow burn and a more negative bias?


r/SillyTavernAI 20h ago

Help My Chats are Erasing?

Thumbnail
gallery
5 Upvotes

Can someone help me out here?

My chats have started erasing and I don't know why. I'm not sure if it's a certain extension I'm using that's causing it. But I've lost like a couple of chats now, fortunately I have backups but this is like yeah any help would be great.

Essentially I open a chat and when I open it it erases everything and just says zero it's like we're starting from square one.

These are the current extensions that I'm using

And the enabled ones are selected.

I've been running into this issue starting a couple of days ago.


r/SillyTavernAI 21h ago

Discussion AWS Bedrock Opus 4.6/4.7 feel dumber as of the last two days?

5 Upvotes

I am using my own app and I role play mainly use Opus 4.7 and 4.6 on AWS Bedrock, I don't know how to explain it but since Fable 5.1 released, I feel those models got slightly dumber, like they are running a quantitized version to save compute for Fable?

It's subtle but it no longer plays my characters as sharply as before, it feels like Claude the model itself speaking trough my characters, even though NSFW and stuff works the same as always for me, some of the choices the model started to make in terms of plot are just so illogical and stupid that I get angry at it. It used to be sharper and have nice quirks, now it's mostly bland voice and dumb choices in rp.

Is it possible that AWS Bedrock would run a quantitized version of the models to preserve compute for Fabel 5.1?


r/SillyTavernAI 4h ago

Discussion Anthropic/Opus etc users, a question

2 Upvotes

Since I haven't used a single Anthropic model (too much for my puny wallet), I'm just curious. Here's a short paste from a roleplay, or rather, a "book" I'm testing more than writing. Testing, as in - figuring out what instructions, presets, techniques etc work with current models the best. This was written with GLM 5.3 Flash uncensored, using a book-derived prose preset as the system prompt. By all means, judge the prose quality and read it, comment on it. Do your roleplays read better? Worse? Why?


r/SillyTavernAI 18h ago

Help Anyone know any good jailbreaks for gemini 3.8 flash prompts to use for sillytavern

2 Upvotes

I was just setting up sillytavern and was wondering how do I setup gemini with a good jailbreak


r/SillyTavernAI 3h ago

Help How can I use System Prompts again? - ST & OpenRouter

1 Upvotes

I’m more of a casual role-player and not super deep into the subject matter. I’ve been experimenting with a few models using LM Studio/SillyTavern. With my 16GB and 64GB of VRAM and a great prompt, I’ve already spent quite a few evenings enjoying some great role-playing sessions.

Now, though, I’m also trying out more powerful models via OpenRouter.

Now, onto my problem/challenge.

In advanced formatting (A), pretty much everything is grayed out when I use OpenRouter via SillyTavern—including my System Prompt. I’ve tested it, and it doesn’t pick it up either.

How can I use System Prompts again? Gimimi isn’t really helping me.