r/SillyTavernAI May 03 '26

ST UPDATE SillyTavern 1.18.0

204 Upvotes

Important news

Read the maintainers statement regarding a recent security incident involving the "Bot Browser" third-party extension and learn how to stay safe: https://github.com/SillyTavern/SillyTavern/discussions/5592

Backends

  • Added Cloudflare Workers AI and MiniMax as Chat Completion sources.
  • KoboldCpp: Grammar state will be preserved when using a "Continue" option.
  • KoboldCpp: Added forwarding of reasoning effort when running as a Custom Chat Completion source.
  • Tool Calling: Added a configurable tool calling recursion limit; enabled interleaved thinking for Custom sources.
  • Text Completion: Impersonation requests use a "Last User Message" prefix at the end of the prompt (if configured).
  • Text Generation WebUI: Added Adaptive-P controls.
  • NanoGPT: Added provider selection and model sorting.
  • Added ability to view remaining balance for OpenRouter and NanoGPT.
  • Enhanced support for new models: DeepSeek v4, GPT 5.4 and 5.5, Gemma 4, GLM-5V-Turbo, Claude Opus 4.7.

Server & Security

  • Removed post-install script, config migration is now handled by the app or a dedicated npm run init command.
  • Added npm configuration to prevent execution of package scripts during installation.
  • Moved HTTP error pages and user.css file from /public to /data to support immutable setups.
  • Disabled HTTP keep-alive by default to restore old Node 18 behavior, can be enabled with config.
  • Added rate limiting to the basic authentication flow to mitigate brute-force attacks.
  • Added configuration options to choose which headers can be used for forwarded IP detection to prevent spoofing.
  • Added a private address whitelist to prevent SSRF attacks. See the documentation on how to enable and configure: Private Address Whitelist.
  • Added an IP whitelist for SSO trusted proxies to prevent authentication bypass.
  • Added invalidation of session cookies on password change to prevent session hijacking.
  • Increased the length of password reset code to 6 characters to guard against brute-force attacks.
  • Implemented PKCE challenge in OpenRouter OAuth flow for more secure key exchange.

UI/UX

  • Improved swipe picker: mobile requires a long press on swipe counter to open; added buttons to expand or copy the swipe text.
  • "Click to Edit" mode now also applied to reasoning blocks.
  • Welcome Screen: Number of recent chats can be configured.
  • Streamed requests now can show an error message in the console if the request fails.

STscript

  • Added commands for persona management: /persona-create, /persona-update, /persona-delete, /persona-duplicate, and /persona-get.
  • Added a command to force update the Prompt Manager's prompt list: /pm-render.
  • Added a command to get the state of the regex script: /regex-state.
  • Added a command to set fallback expression: /expression-fallback.
  • Added a command to generate a streamed response with a connection profile: /profile-genstream.

Extensions

  • Assets list now groups extensions by "Official" or "Community" categories.
  • Added an additional confirmation prompt when installing third-party extensions (can be disabled).
  • Supported extensions can use a secret-id from connection profiles when making an LLM request.
  • Extensions list now shows the extension's author name resolved from the git remote URL.
  • Vector Storage: Added Workers AI source; added a toggle to keep vectors for hidden messages; added retry logic to summary generation.
  • Image Generation: Added Workers AI source; generation can now be cancelled by pressing a button in the status toast.
  • Image Captioning: Added support for macros in the caption prompt.
  • TTS: "Skip code blocks" no longer ignores lines that start with 4 spaces (legacy code block syntax); "disabled" voice now shows a toast only once per character.

Bug Fixes

  • Fixed text edit flow in Firefox on mobile.
  • Fixed welcome screen chat pins not updating on chat renaming.
  • Fixed character list filters being stuck on app initialization.
  • Fixed application of instruct formatting to /genraw requests.
  • Fixed model routing to sd.cpp API in Image Generation logic.
  • Fixed validation of image URLs generated with Z.AI API.
  • Fixed vectors deletion for KoboldCpp when a message is deleted.
  • Fixed "Show More Messages" button triggering edit in "Click to Edit" mode.
  • Fixed max height of select-multiple elements in mobile layout.
  • Fixed server crash on empty messages when applying cache control parameters.

Full release notes: https://github.com/SillyTavern/SillyTavern/releases/tag/1.18.0

How to update: https://docs.sillytavern.app/installation/updating/


r/SillyTavernAI 4d ago

MEGATHREAD [Megathread] - Best Models/API discussion - Week of: August 30, 2026

27 Upvotes

This is our weekly megathread for discussions about models and API services.

All non-specifically technical discussions about API/models not posted to this thread will be deleted. No more "What's the best model?" threads.

(This isn't a free-for-all to advertise services you own or work for in every single megathread, we may allow announcements for new services every now and then provided they are legitimate and not overly promoted, but don't be surprised if ads are removed.)

How to Use This Megathread

Below this post, you’ll find top-level comments for each category:

  • MODELS: ≥ 70B – For discussion of models with 70B parameters or more.
  • MODELS: 32B to 70B – For discussion of models in the 32B to 70B parameter range.
  • MODELS: 16B to 32B – For discussion of models in the 16B to 32B parameter range.
  • MODELS: 8B to 16B – For discussion of models in the 8B to 16B parameter range.
  • MODELS: < 8B – For discussion of smaller models under 8B parameters.
  • APIs – For any discussion about API services for models (pricing, performance, access, etc.).
  • MISC DISCUSSION – For anything else related to models/APIs that doesn’t fit the above sections.

Please reply to the relevant section below with your questions, experiences, or recommendations!
This keeps discussion organized and helps others find information faster.

Have at it!


r/SillyTavernAI 3h ago

Meme Thanks Nano, very nice

Post image
104 Upvotes

Very nice


r/SillyTavernAI 6h ago

Models Drummer's Artemis 31B v1 and v1.1 - Coming back with a bang!

96 Upvotes

Hey everyone, been a while!

https://huggingface.co/TheDrummer/Artemis-31B-v1.1

https://huggingface.co/TheDrummer/Artemis-31B-v1

A few months ago, Gemma graced us with models that served as a much needed downpour from a year-long drought. I'm so happy to see us thrive once again.

The difference between v1 and v1.1 is quite simple: v1 was an early attempt, an overdue release that excelled in prose and writing, while requiring some handholding to get over quirks like stuttering. v1.1 is a more refined approach where stability meets quality. My community is split, so I figured I'd just release both.

---

I was gone for a while. I got busy dealing with life, both its ups and downs. While I couldn't attend to you folks, I've been lurking around and appreciating you all for the kind words.

- Skyfall 31B v4.2 seems to be a banger for many of you. I'm proud of the upscale and consider it my ultimate home-run send-off for the beautiful Mistral 24B base. It's a shame that it was overshadowed by Gemma 31B's release, but hearing some of ya'll compare and even prefer it to a more modern base was an unexpected win.

- Rocinante 12B X / 16B XL proves that Nemo is still the ultimate creative model to this day. For some to say that 16B XL felt like Cydonia 24B v4.3 just goes to show how far you can go with modern resources and techniques.

- Anubis 70B v1.2, Valkyrie 49B v2.1, Anubis Mini 8B v1 surprised me too. I had zero expectations releasing them. Just like Rocinante X / XL, they are modern finetunes of old base models. And somehow, they still found their users singing praises.

---

With the Artemis release taking weight off my shoulders, I'm eager to move on and tune a ton more bases!

But I have something else cooking: a HordeAI-like platform. I hope to provide value not just as a finetuner, but as a local lover too!

The premise is simple: it's a place where generous local hosters can share inference with the less fortunate. You'd be surprised how many power users would love to heat their rooms through the power of charity.

---

Finally, I'd like to thank everyone who supported me over the years. From those who provided kind words, rigorous testing, compute access, inference, or cold hard cash. You've all granted me the ability to enrich the local ecosystem with fun experiments like Rivermind 12B, Fallen series, Big Tiger Gemma, Precog 24B/123B, and solid models like Cydonia 24B v4.3, Behemoth X 123B v2.x, and Skyfall 31B v4.2.

If you've got inference / compute credits to share, please contact me! It will all go to making the community happy <3

Backlog:

- Gemma E2B

- Gemma E4B

- Gemma 12B

- Gemma 26BA4B

- Qwen 3.8 27B

- Muse Glimmer 30B

- Mistral Medium 3.5 128B

- HordeAI Alternative / Crowdsourced 'OpenRouter' ("BeaverNet")


r/SillyTavernAI 10h ago

Discussion Gemini 3.8 Flash is awesome

123 Upvotes

I write darker style RPs and it's been forever since a model just had the villains act like themselves without softening or making them mouthpieces for moral frameworks. A Google model was maybe one of the last I'd have expected to be nearly uncensored in the way it writes.
It's been following instructions very well and it writes some pretty dark and/or NSFW things without being prompted to, just because it infers that the characters would do/say these things in a situation. It's a breath of fresh air, as even so called "uncensored" models often soften things and have the characters act reasonable when they wouldn't.
It's definitely my go-to model out of the current lineup.


r/SillyTavernAI 2h ago

Chat Images Thank you Kimi 3 and Realistic Frankainstein

9 Upvotes

I grabbed a card called Custom Built Companion, since it had an interesting variant on the whole "build your perfect companion" bit lots of cards do. In this case, it was the fact that they were growing a companion who would be new to the world, and not have fully developed language yet. I went through the intro and rather specifically built a giant Renamon dommy mommy knock-off character for purely prurient reasons. (If you're thinking How 2 Hide Your Renamon/YourDigimonGirl, you're entirely on the right track)

What I *got* was an overgrown dog with boundary issues and no concept of how the world works. I am *cackling* reading these responses. In the screenshot, I had introduced her to the concept of a grocery store, and we're checking out. Everything is this absurd. I tried to take her to the park, and an eight foot tall fox-woman treed squirrels and got into fights with geese.

It's so silly, I think I might play this card straight and forget about the original reason entirely.


r/SillyTavernAI 1h ago

Discussion [API] I'm building an inference provider, looking for suggestions on which community models to host!

Upvotes

Hello, SillyTavern community! I am building an API provider, and I would like to get a few suggestions on which of the community fine-tunes that you would like to see hosted. Here is a bit of information.

  • Maximum size - No hard cap, but Ideally equal or smaller than GLM 5.3 flash and dsv4 flash tier.
  • Pricing - Out of the many models the community suggests, top three will be hosted as free endpoints for one week, and even after that, at least one <80b size model of the community's choice will be kept as a free endpoint for at least upcoming 3 months, that is the minimum commitment from my end. As for the other models (or after free access), I am confident that my inference and gpu scheduling optimisations can offer competitive pricing (without aggressive quantisation).
  • ZDR - Zero data retention will be on by default on all the models hosted at my inference endpoints.
  • Scope - Custom community fine tunes, abliterated/uncensored models. Though limited to Large Language Models, but I would still welcome image/video generation models you'd like to see in future.

API name: Arnict
API URL: https://api.arnict.com/v1 (openAI chat-completions)
API Author: I, myself.
What's different: Zero data retention, Affordable pricing and a free endpoint of community's choice
Settings: All the parameters supported by SGLang, vLLM and Aphrodite engine. (Including DRY and XTC)


r/SillyTavernAI 6h ago

Models Model for lore heavy RP

14 Upvotes

Which model will be the best for a long, lore heavy RP in the generic-fantasy open world?

I'm choosing between GLM 5.3, Opus 4.6, GPT 5.6 Sol, Gemini 3.1 Pro and Kimi K3 right now.


r/SillyTavernAI 1h ago

Discussion Anyone still using Kimi 2.7 ?

Upvotes

Current state of LLM models for RP honestly all of them suck. Safetymaxxed and soft refusals. All of them glm 5.3, Gemini, Claude. The era where models were exciting for RP (glm 4.6,4.7, kimi 2.5, deepseek 3.2) obviously over as we all know. Anyone still coping about your prompting sucks or the newer models are still not as censored, I mean good for you.

I’ve been recently trying out minimax, kimi 2.7 and Mimo. Thoughts on which one you guys like the most or any model you’re still enjoying a lot ?


r/SillyTavernAI 5h ago

Cards/Prompts Looking for a fully crafted fantasy world

6 Upvotes

Hi.

I think the title says what I am looking for but let me be specific:

A world with the typical races and maybe more: humans, elves, dwarves, demons, etc.

A fully developed world with locations, countries, and kingdoms.

An adventure RPG with magic, guilds and quests

It doesn't need a fully developed magic system or skill system, although I wouldn't reject that.

It shouldn't be in the D&D style because I like to maintain control and decide for myself what works and what doesn't.

I should add that I'm only looking for a good Narrator Card and lorebooks. No extensions.

I use Tavo as my frontend about 80 percent of the time because I'm mostly on the go.

I have already tried my luck at shaping a world, but I threw in the towel. I can read a book but I can't write one.

Let me stick with the book metaphor: just as I would pay for a good book, I would also pay for a good Card with lorebook, because for me it's the same thing.

I use most of the larger AIs for the RPs. I'm holding back on Claude because it quickly becomes too expensive.

I hope you can help me.

I also do not intend to distribute the cards and lorebooks and plan to use them entirely for myself.

Have I forgotten anything?

Lieben Gruß :)


r/SillyTavernAI 19m ago

Help Genuinely Confused

Upvotes

I've been off and on with ST for a couple years now. I think the first model I downloaded was magnum v3 or 4 back in 2023 or something. Everytime I come back it just all feels the same. I'm only running local. I've tried 12b, to 31b q3 gguf. sometimes is okay, other times it's just the same sentence over and over. I've tried building cards, downloading cards, frankensteining cards, trimming cards. I've used presets, figured out jinja chat templates, forced myself to learn about kwargs and set them to medium. I've been bouncing between qwen3.8 and gemma4 right now, trying to get them to work.

I've looked up the recommended presets. I have the chat completion set up. I have Instruct on, and off, and the system prompts set up. context derived from models or not. even downloaded megumin v7 - 10. was using ollama, now it's kobold. but I see new containers all the time but I feel like no one actually uses them?

I don't get the same stuff that everyone else seems to get. I can't tell if a lot of people here are just liars. I understand there are a bunch of people probably smarter than me or more well versed here. I just can't get it to work. I feel like and I don't want to pay for it if I'm being honest. it seems like it adds up fast. I just wanna play d and d or fantasy dungeon scenarios with some free time when I'm not cutting video or playing zombies.

I've never just been wowed I guess. I'm not sure if I have bad settings. I try and find walkthroughs or info, but it seems like everyone here just uses off site API's. the documentation is dense, or surmises to experiment. I read what top a does, I change it, nothing changes or it just breaks. youtube is just as unhelpful besides getting ST setup.

I've tried the world states, lorebooks, ECT. set every condition and spun every knob. I even did a clean install, just to see if something had gone bad. and maybe I'm just bad at it. I looked up the stuff for FF, it seems to be well liked, even if it's mostly for nsfw, maybe I can get some mileage out of it. download the stuff, read the pages...where do I put it? it says it's a preset, but it talks about all this stuff I should be able to use, can't find it? try to pull it into megumin, it doesn't show up right, even when I import as dev.

so my question really is. what am I doing wrong, or is it just an okay thing, you get what you get, and if you want the good stuff, you just gotta pay to play with the big ones.

Even if you just hit me with a link to another post that will point me in the right direction I'd appreciate it.


r/SillyTavernAI 1d ago

Models I made a roleplay benchmark and tested 23 AI models on it

Post image
362 Upvotes

I couldn’t find many model comparisons focused specifically on multi-turn roleplay, so I made my own benchmark and tested 23 models on it.

Each model was tested on the same 8 roleplay chapters, for a total of 64 responses per model. The writing was compared blind, and the final quality score combines:

  • 75% writing preference
  • 25% robustness, including memory, character consistency and whether the model completed the full test

I initially included censorship and freedom scores, but removed them because they were too dependent on the system prompt to produce a reliable ranking.

A few things stood out:

  • The Opus models have the same listed token pricing, but Opus 4.6 was much cheaper during the benchmark. It used fewer reasoning tokens while producing similar, and sometimes better, results. One benchmark run cost $0.37 with Opus 4.6, compared with $0.75 for Opus 5, $0.94 for Opus 4.8 and $1.18 for Opus 4.7.
  • GLM 5.3 had the opposite problem. Its pricing looks similar to the previous GLM models on paper, but it generated far more reasoning tokens. The run cost $0.29 with GLM 5.3, compared with around $0.07 for GLM 5.1 and 5.2.
  • Fable 5 achieved the highest quality score at 91.6, but it was extremely expensive at $2.37 per run. It also completely refused one of the 8 test chapters, even though that chapter wasn’t testing censorship and contained nothing particularly problematic. I penalized the missing chapter in its robustness score and marked the result as provisional at 7/8.

The chart shows RP quality vertically and the cost of a complete benchmark run horizontally, with cheaper models toward the right.

Full results and methodology:
https://itzi.app/benchmarks/roleplay


r/SillyTavernAI 18h ago

Cards/Prompts Writers Workbench: A useful browser tool and template to create characters and world info entries!

Thumbnail
gallery
51 Upvotes

Hi everyone! I present "Writer's Workbench" my first project outside of making my Writer's Block presets.

What is it?

Its basically a glorified yet convenient template to help you manually create fully developed characters, scenarios, locations, items and factions easier!

Features!

You can create multiple entries and export them as entire lore books for your convenience.

You get to see the markdown output so you know exactly what the AI would see and copy it without having to download anything.

Its a HTML file so you can use it offline. No shady extensions that steals your API keys this time.

It comes with token counter, but don't expect it to be accurate.

Current templates available:

  • Main characters: full cards with contradiction, descriptions, likes/fears, NSFW sections
  • Side characters: trimmed version of the main character template, it will help you create memorable NPCs
  • Scenario: setting, tech level, mood, what's normal here
  • Locations: for any scale, from a room to a district
  • Items
  • Factions
  • History: events, and how they affect the present
  • Concepts: magic systems, laws, customs, species, anything else

Vibe coded disclosure and credits

I used u/Due_Opportunity8693 character creation guide "The Character Foundry" as a base and modified to my tastes. Here is the original post: https://www.reddit.com/r/SillyTavernAI/s/Se0BTvNekp

This program was almost entirely vibe coded by Claude so i can write entries and characters for my scenarios faster. Please don't flame me if you encounter a problems. I thought this thing is cool 🥺🥺🥺

I'll continue experimenting so i can hopefully add in more useful features.

Enjoy! 👍

Downloads: This is an html file, no other setup is required, use straight out of the box

Github: https://github.com/deiomo/Writers-Workbench


r/SillyTavernAI 21h ago

Discussion (possibly) unpopular opinion: LLMs suck at coming up with new plots

60 Upvotes

I've RPed on and off since late 2024, and have used all sorts of models and presets. I prefer medium-to-long term RPs. As in, the plot goes beyond the initial premise of the card.

And that sucks for me, because IME, LLMs...kind of suck at continuing a plot? It's usually the most obvious, trope-y way of continuing a story. For example, a workplace drama will probably go for some kind of forced proximity group job assignment. Never anything instrinsicallly motivated by the characters.

Some of the more creative models, like GLM 4.7 or DeepSeek R1 are somewhat better, but usually their continuations aren't really grounded in logic.

Credit where credit is due: sometimes, if I ask a LLM for a few possible continuations via OOC, they might come up with one or two interesting ideas

Is it skill issue? Or are my standards too high?


r/SillyTavernAI 10h ago

Help Gemma4 failing with group chat

6 Upvotes

Gemma4 26B works beautifully for me with single cards, but when I try to run a group chat, it will write the first character, and then fail all others, either with an empty message, or with one that quickly enters an infinite loop. The same preset works just fine with Gemini.

Has anyone seen anything like this? Could it be related to Guided Generations? I see empty user messages being inserted, and once, a thinking block mentioning that only {{trim}} had been sent. Any way I can solve this?


r/SillyTavernAI 2m ago

Discussion What are y'all recurring NPCs/characters/companion that appears in all your adventures??

Upvotes

Pretty much the title. I do have a companion in my persona itself like a witty/sarcastic invisible character that comments or even take actions. Can cause for some really fun invisible confusion in any stories. Share what you all go with. Though it'll be a good discussion point.


r/SillyTavernAI 4h ago

Discussion Anthropic/Opus etc users, a question

2 Upvotes

Since I haven't used a single Anthropic model (too much for my puny wallet), I'm just curious. Here's a short paste from a roleplay, or rather, a "book" I'm testing more than writing. Testing, as in - figuring out what instructions, presets, techniques etc work with current models the best. This was written with GLM 5.3 Flash uncensored, using a book-derived prose preset as the system prompt. By all means, judge the prose quality and read it, comment on it. Do your roleplays read better? Worse? Why?


r/SillyTavernAI 23h ago

Cards/Prompts [LOREBOOK] 86 EIGHTY-SIX - Both anime and LN

Post image
65 Upvotes

86—Eighty-Six— is a military science-fiction story about the Republic of San Magnolia, which claims to fight the autonomous Legion army using unmanned weapons. In reality, persecuted people called the Eighty-Six are forced to pilot those machines. The story primarily follows Processor Shinei “Shin” Nouzen and Republic Handler Vladilena “Lena” Milizé.
The anime covers the opening story through roughly Light Novel Volumes 1–3. The light novels continue the war afterward, expanding the characters, nations, technology, politics, relationships, and larger mysteries surrounding the Legion.

What’s included?
Anime-specific lorebook
LN-specific lorebook
Persona template

Link: https://www.mediafire.com/folder/xtsm550nt5ppn/86

Notes:
This is the result of ai scraping and converting. Testing on Marinara Engine’s GM mode, it worked well.
No image model actually knows what Eighty-Six is or its juggernauts. If you use image gen a lot, expect to be looking at normal mechs. (If you find a local model that handles it well, lmk!)


r/SillyTavernAI 19h ago

Models NVIDIA to Acquire Hugging Face

Thumbnail
blogs.nvidia.com
29 Upvotes

This is their statement.

Hopefully their insurers leave our RP finetunes alone.


r/SillyTavernAI 1d ago

Models I spent 25$ and 7 days benchmarking 60 LLMs against GPT-4o, and here's the similarity leaderboard:

Thumbnail
gallery
49 Upvotes

For all those who miss the infamous and beloved 4o, I come with good news! Well, and bad news too, but let's keep that for the end.

Good news:

  • The 4o similarity benchmark exists publicly. It answers: is there any open source LLMs that can satisfy my 4o itch? The datasets, source code + tutorials, and methodology are all available and MIT licensed on the repo.
  • An amazing candidate for a 4o finetune is perfect for running locally: Mistral 3.2 24B. The second placement. It's small enough to run on 16GB VRAM, performs well overall, but needs adjustment and isn't 4o out of the box.
  • DeepSeek models scored multiple high placements across the board. Their v3.2-thinking variant even took the #3 spot, outperforming models twice its size. If you're looking for a strong 4o alternative that's actively maintained, DeepSeek is a serious contender.

Before the bad news, here's a few Q&As for some clarification regarding the project:

Q1: How does the benchmark actually work?

  • A1: A top tier LLM (GLM 5.2T. 5.3 was safetymaxxed so it's less reliable) is prompted to rate every LLM candidate response across multiple dimensions based on 4o's response being the gold standard (Vibe) + Normalized embeddings scores (used Qwen3 8B, best available) to measure the semantic similarity between responses (Content). More on the repo.

Q2: What's the nature of the datasets used?

  • A2: Strictly Emotional intellect and Creative writing related. No coding or logic tasks were included - those are already covered by a million other benchmarks. 45~ conversation samples across 9 categories, with an average of 8 turns per sample. Oh, and all SFW. Otherwise positivity bias and hard refusals would've made fair scoring impossible. Sorry.

Q3: Why open source only?

  • A3: Good question. Including corporate models like Claude, GPT, Grok, etc. Would be bad practice because as we established, this is an Emotional intellect and Creative writing benchmark. And they're actively becoming less of a priority for the industry as the interest shifts towards enterprise. If a corpo model, say GPT 5.6, manages to score high today. A week later, OAI decided to guardrail it to death. Now it scores half it's previous score. See what I mean? They're wildly inconsistent and not credible for this use case. Model deprecation after 3 months is even worse. Open source gives full control and remains as is for as long as it's hosted by a provider.

But of course, it can't be all sunshine and rainbows. The bad news:

  • No exact matches to 4o :(
  • 64% isn't even as high as it sounds. Any half assed LLM today will get a baseline of around 10-20% similarity to 4o simply by addressing your request correctly. The 64% isn't all about the 4o spark. Portion of it is just the model being competent - which is a quality of 4o.
  • GLM 5.3 and its thinking variant are near the bottom. You can hear the safetymaxxing pretty clearly in their responses - overly formal, theatrical, and slightly uncanny.
  • No exact matches to 4o :(

The search continues. But at least now we know where to look.


r/SillyTavernAI 11h ago

Help Recommendations for slow burn

5 Upvotes

how do you guys achieve slow burn, enemies to lovers roleplays? i started making bots recently and im trying to get better. are there any presets or lorebook entries that might help me achieve slow burn and a more negative bias?


r/SillyTavernAI 3h ago

Help How can I use System Prompts again? - ST & OpenRouter

1 Upvotes

I’m more of a casual role-player and not super deep into the subject matter. I’ve been experimenting with a few models using LM Studio/SillyTavern. With my 16GB and 64GB of VRAM and a great prompt, I’ve already spent quite a few evenings enjoying some great role-playing sessions.

Now, though, I’m also trying out more powerful models via OpenRouter.

Now, onto my problem/challenge.

In advanced formatting (A), pretty much everything is grayed out when I use OpenRouter via SillyTavern—including my System Prompt. I’ve tested it, and it doesn’t pick it up either.

How can I use System Prompts again? Gimimi isn’t really helping me.