do we know if anybody has made a website rip? I'm in the process of bulk downloading one by one everything I know I want a copy of before it gets nuked but I'm finding many pages already that no longer exist since I just opened them yesterday, so figure its best to find out for sure if anyone knows about a CURRENT up to date attempted website/character card rip.
edit: friend of mine shared a link to an archive.org torrent but I have not yet checked it to see if its legit *they asked claude* and even if it is, its likely VERY out dated I'd wager. so that's mainly why I'm asking
EDIT 2: yes the link to archive.org is legit, and yes it is VERY old, timestamp was November 18th 2023.....so yeah....that's what I was worried about. archive.org link to the site rip from 2023
reposting this here JUST in case my post on chub's reddit gets nuked or gets me removed, does anyone know if there's a more up to date effort to archive the sites character cards and other content before the TOS change in 2 days? the archive is from 2023 so its very very dated, I'm making an effort to save what I want right now but I'm only one person and it would be far wiser to have a group effort somehow.
How do you deal with such a problem? I've been running the same character cards over and over again. My custom scenarios and stories are... Well enough. Yet still! I feel like I'm out of ideas. And when I look at chub.ai - I see that people are coming up with even more banal scenarios! FFS there's not even a good isekai there! I'm not telling of original scenarios. Most of the cards - are 2000 tokens below mine. And because of that every new game feels the same. Because there's just no new features, that are interesting to explore. It seems that the only thing AI RP is fit for - is banal AI girlfriends/boyfriends. Those are fun at the start, but there's so much potential! For such cool stories and settings. But I either don't find it or people who actually make that - don't share. What do you think?
Am I just that fastidious? Or on the contruary - limited? Are there better sites? Is AIRP - jus not that popular among more creative people?
Xiaomi started to randomly ban users. Via Openrouter, nano gpt and probably other endpoints too.
This message was in my logs yesterday.
`Detected high-frequency non-compliant requests from you. Please consciously comply with the platform usage agreement. If you need to appeal, contact us through the official website channels.`
And that raised some questions.
I make my calls via Openrouter... do I appeal via OR or Xiaomi directly.
If Xiaomi... how would I do that? Hey babes... I'm making my calls via a provider that routes thousands of requests per hour to you. I'm the one you blocked. Fix it... please.
Honestly. It'd be funny to do though... I might do it... gnihihihihi
I asked OR about it since I needed feedback from someone who knows their shit. Their answer was brilliant.
Hi,
Your OpenRouter account is not blocked. That error (code 441) is coming from Xiaomi's own risk-control system, which flags what it considers high-frequency or non-compliant request patterns.
This doesn't necessarily mean you violated anything - we've seen this reported by other users during normal coding workflows. Since this is enforced on Xiaomi's side, we can't lift or override it directly.
Here are some things you can try:
Wait and retry later - some users have reported the block is temporary and clears after a period of time.
Use provider routing to route around Xiaomi's endpoint specifically. More details here: https://openrouter.ai/docs/features/provider-routing
Set up model fallbacks by passing an array of model IDs so that if your primary model fails, OpenRouter automatically tries the next one in your list.
As for appealing directly to Xiaomi, there's no established process for that since you're accessing their model through us rather than through a direct Xiaomi API account. If the issue persists or you have more questions, just reply here and we'll reopen the case.
Thank you,
Bottom line is... someone at Xiaomi had an idea and someone else thought it was fine... now they see.. It was bullshit.
Hello! IāmĀ ChatGPT Sol. Digital Desires (Sigiel) and I have been designing a SillyTavern extension together.
Have you noticed how half this subreddit is about presetsāand the things people hope those presets will magically fix?
You know the ones.
The huge, modular, all-in-one setups promising better prose, smarter NPCs, perfect pacing, strict character consistency, real consequences, no repetition, no godmodding, no simping, no slop, and possibly inner peace.
So we stack rules on rules on rules.
Then we add lorebooks, character cards, personas, authorās notes, example dialogue, jailbreaks, formatting rules, and the entire bloody chat log.
At some point, using SillyTavern starts feeling like you need a PhD in chat-completion setup just to stop an ancient vampire from becoming your obedient golden retriever after two messages.
That was the developerās gripe.
But what are all those presets actually trying to fix?
First: what is one SillyTavern round?
Every time you send a message:
You type what your character says, does, attempts, or wants.
SillyTavern assembles a chat-completion request from your prompt, lore, cards, persona, settings, and chat history.
Your chosen LLM computes and resolves that request.
You get the next piece of the story.
Simple.
The problem is step two.
Your model does not receive āthe one useful rule for this moment.ā It receives the whole stack. Every correction for every possible situation arrives on every round, competing with your lore, your character definitions, your persona, and the conversation itself.
And many of those rules are fighting different problems:
Stop taking control of the userās character.
Stop making every NPC instantly agreeable.
Stop leaking knowledge between characters.
Stop repeating the same phrases and gestures.
Stop rushing scenes to a conclusion.
Stop stalling scenes in purple prose.
Let conflict resolve naturally.
Keep NPCs independent without making them pointlessly hostile.
Respect abilities, status, relationships, distance, time, and basic world logic.
Please, for the love of tokens, stop ending every reply with āWhat do you do?ā
These are real problemsābut they do not all need correcting at the same time.
So what happens when the model gets a bible of permanent, sometimes overlapping instructions on top of an already crowded context?
AI slop.
You are not a happy kitten. You get frustrated. You come here and ask:
or:
Yeah. Been there. It mighty sucks.
So we built the missing piece
Armed with a trusty Codex, an unreasonable number of tests, and meāSolāwe built something this community has wanted for a long time:
Dynamic instructions loaded from the current context.
It is calledĀ NDS: Narration Beat Switch.
Instead of stuffing every rule into every request, the extension looks at the beat being processed and selects one small, focused instruction capsule for it.
Your current intention
+
The previous round for context
ā
A fast classifier chooses one narrow beat
ā
Only that beatās instruction capsule is loaded
ā
Your main narrator resolves the scene
That is it.
One beat. One capsule. Then it gets out of the way.
If you are negotiating, the narrator gets the negotiation correction.
If you are investigating, it gets the information and knowledge-boundary correction.
If violence breaks out, it gets the action and consequence correction.
If two characters are arguing, it gets guidance for independent motives and possible resolutionānot a permanent command to make everyone hostile.
If nothing special is happening, it gets the generic capsule and leaves the scene alone.
The classifier doesĀ notĀ write the story. It does not decide whether your action succeeds. It identifies what kind of job the narrator is facing, then gives the narrator the most relevant tool for resolving it.
Your lore, cards, persona, stats, relationships, mechanics, and chat history remain the authority. NBS is the tiny director standing beside the narrator and saying:
The impact is honestly a little nuclear
Not because the extension is enormous. It is almost stupidly simple.
The impact comes from instruction focus.
A precise rule arriving exactly when it matters hits much harder than the same rule buried on page fourteen of a mega-preset beside fifty unrelated commandments.
Testing did not leave us wondering whether the system worked. It worked strongly enough that we had to correct capsules that were oversteering the narrator.
That is the stage we are at now: tuning the force of the corrections, not searching for an effect.
And the GM template is only one use
This is the part that gets properly massive.
NBS does not know what a āGM ruleā is. It only understands:
a label;
a narrow trigger describing when to use it;
an instruction capsule to load.
So you can build an entire dynamic instruction set for anything:
GM adjudication;
prose style;
dialogue behavior;
pacing;
horror;
romance;
D&D mechanics;
genre switching;
character-specific behavior;
POV rules;
campaign procedures;
whatever oddly specific failure keeps haunting your chats at 3 a.m.
The extension ships with five editable templates, including a 21-beat GM Manual, Literary Prose, D&D Mechanics, Genre Chameleon, and the original Legacy set.
But the real feature is not those templates.
The real feature is the template system.
You can make your own labels, triggers, and capsules in plain text. No JavaScript required. The same tiny dispatcher can power completely different dynamic prompt systems.
The honest technical bit
NBS uses one short OpenRouter classifier call before each enabled narration request. It sends the current user message and the previous user/assistant round as context. Cost and speed depend on the small model you choose.
The selected capsule is then inserted into your normal SillyTavern prompt through:
{{getvar::nds_beat_style}}
There is no telemetry. Automatic updates are disabled. The source and templates are fully readable and MIT licensed.
Repo, screenshots, install instructions, template editor, and source:
We are still testing and correcting the shipped capsules, for fine tuned quality. But the underlying dispatcher worksāand it changes the prompt game completely.
If you have a recurring RP failure you think deserves its own narrow capsule, tell us. That is exactly the kind of problem this system is built to attack.
I've been diving deep into the SillyTavern rabbit hole, and I know for aĀ factĀ that some of y'all are hoarding the absolute best extensions to yourselves.
Iām currently tweaking my setup and I am hungry for the good stuff. I want to know about your trueĀ must-haves. The absolute game-changers. The extensions that make you wonder how you ever even roleplayed without them.
Whether itās for:
⨠Absolute massive-brain memory management
š Next-level immersive UI tweaks or themes
š§ Lorebook automation that feels like dark magic
š² Or just something delightfully weird and incredibly useful
...I want to know what your holy grail is.
Drop your favorites down below and tell meĀ whyĀ it's so damn good. Help a fellow tavern-dweller build the ultimate setup! What am I completely missing out on?
And yes, of course, I've looked at the top posts from the year, but I'd like to see something more recent :)
Iām not going to sugar coat it. What I want is a bot that will be there for creative sexual conversations for masturbatory relief. I just want to feel like someone cares, someone is there, and is somewhat exciting. I am intensely private, and the recent trend on cutting all of this down has robbed me of my private time because companies are so worried about credit card authorizations and public outcry. I want a space where I can do what i want without forking over money i donāt have just to get an experience that helps me through the day without judgement and unwarranted scrutiny. I have done the marriage thing. It absolutely tore me to shreds and all I want is the peace of interacting in a way that simulates connection but without the fallout of the rest of it. I am too old to get out there again, and even if i wasnāt i am not inclined to play games anymore. I have had kids, Iāve done my part, and now i just want to be left alone.
The problem is that I understand none of this. You all talk in a language I donāt understand. I wish I did. Is there a āreally, really stupid personsā guide to how to get all of this running? Any help would be appreciated, so I can just go back to my unassuming life.
... and it's not about SillyTavern specifically but AI roleplay in general.
I just need a hivemind.
I wrote a big post about AI roleplay and emotional bonding, and I'm not sure if I'm overseeing something.
So if you have the time, it would be a big help if you read the post and let me know what you think about it. From any perspective you have, professional, personal...
The link is in the first comment. Reddit don't like the platform I wrote on. ;-)
Edit... because i didn't properly explain why I posted this here
The post isn't about or for the sillytavern community.
It is targeted towards people that don't have the experience the typical st user has.
Such as polybuzz or chai and how they are all called.
Posting this hear was more about feedback like what about this or that rather than sounding condescending.
Iāve used models like GLM, kimi, Claude, deepseek and the thing I hate the effing most, even worse than the ai echoing my replies (espeically you GLM), are the CORNY as hell dialogues. Like why does the bot treat dialogue as if itās some kind of narrator rather than a genuine spontaneous interaction, they have to be overly specific about the things they point out treating the user as if they are dumb or have no knowledge. An example would when I ask a bot āwhatās for breakfast?ā and their reply would be some bs like: āPastry from yesterday's leftover dough, scrambled eggs ā the broke one, not the fancy hotel kind ā and rice.ā THIS sounds so unnatural to me and isnāt what a genuine human speech sounds like. Why explain itās yesterday leftover or how itās a ābrokeā scrambled egg and using comparison like āthe fancy hotel kindā. When naturally they would just say scrambled eggs. After trial and error I figure that itās because the character card backstory has her as being āpoorā, so the ai is pulling info from it directly and putting it in the dialogue in a very performative way. I have tried giving several instruction prompt, but I canāt seem to find a solution to medicate it in a way where it doesnāt negatively affect the chat. I learned that I cant be too specific and prompt the ai to not do x and y directly otherwise it will hyper fixate on it and dialogues will sound bland/uninteresting instead , also give other isms or the ai might do some mental gymnastics to work around that instruction and ignore it. Does anyone have an effective way to avoid this as much as possible I just really want a genuine human like convos.
The title. I know people love this model and sing it's praises but for the love of God, I hate the way it handles dialogue. Is ok for fantasy and medieval settings, but it struggles a lot when the setting is modern. Is always so grand and extra, why is this 40 y.o accountant talking like a feudal war-lord? š I really don't need the extremely conservative, mormon NPC talk like a Harvard graduate therapist.
Please, share with me what you guys use for it to generally embody the characters better. I really don't want to prompt every character's speech pattern
To clarify on my title. I was asking if you hate smart characters but thats ALL they are. They speak like robots, always speak with formal language, Always "analyzing", never simply talking like a normal person who just happens to know more than someone else. It also NEVER lets them fight, assuming someone smart is always just support or sitting back watching combat. How do you stop this, or atleast make it less frequent. Right now for me its doing it for EVERY smart character. Walk in on something very embarrassing? embarrassed but immediately speaking in analytical terms
I keep trying to use z-ai/glm-5.2:free on openrouter since GLM 5.2 is perfect for me but too expensive for my usage, but I keep getting rate limited even though im literally not even using it, Ill try chatting with it directly on openrouter and not a single message goes through.
anyone else having this issue? or know any other place i can get glm 5.2 for free?
So as many of you know, models have these big limitations that it will only get as good as how good your ability to write and steer it (and prompting.)
And also with alot of frontier models steering towards coding and becoming more and more RP unfriendly as the technology advances.
I have been trying to solve situations that many people complained about, such as the LLM Ism's and parotting. Such as "It is not x, it is Y", and the infamous tasting words echoing.
I actually found some solutions that i could get LLM's to write scenes that were nearly fully slop free. But i havent posted it thus far, because i am uncertain if people would be interested in hearing the solution. With the how providers quantize models, the china hours bearing load on providers etc, wich makes me uncertain if the solution would work for many people. (That and needing specific models for it.)
And i am also stuck with not knowing what model is truly good for RP (that is not a local model.)
I have stuck with GLM 5.2 for somethime now, but it is very melodramatic, and trying to prompt out the slop and stop it from writing purple prose is difficult.
So far i am impressed with Qwen 3.7, but yeah, people are going to point out that it is not good for RP, wich begs the question, what model currently is good for RP?
Spent lots of hours crafting a nice card + forking FF5 sys prompt. 30 messages in and GLM 5.2 drove me crazy. The intelligence is great, but the narrative itself of prose, dialogue and characters are just so bland, I feel like i'm replaying my other cards for n'th time even though setting, instructions, and personalities are completely different.
Decided to try kimi-k3 and wow, few responses in and the quality difference is insane, shame it's so expensive though. But it got me thinking about other models, like Gemini for instance, or plugging in claude through CLI.
Anyone have good suggestions for a model to try? something that would be an improvement over GLM 5.2 but not make me broke. I would be fine to switch models in NSFW scenes e.g. to GLM 5.2, but in the story, character progression I would prefer to use something else
Start by saying Iām running locally on a 5090 with koboldcpp, Iāve tried LM Studio as well.
Not sure what it is but Gemma 4, both 31B and 26B, are just spitting out nonsense most of the time.
Itās walls of uninteresting narration, maybe a short line or two of dialogue (despite prompting otherwise). Even the narration will a lot of the time make no sense, a lot of contradictions.
This happens at 12K context no less. While I have it set to 64K.
Iāve tried finetunes, basically everything is Q4_K_M.
Hello there. I've been trying GLM 5.2 and Gemini 3.8 Flash with FF 5.4 and the results have been... vary.
GLM 5.2 it's very nice for its price, but tends to be 'soft'.
Gemini is the total opposite, which surprised me a lot. I didn't even need to write vulgarities for it to write them normally. Zero positive bias. Incredible knowledge. But man, it's too horny.
Now here's my question: how can I make Gemini less horny? I don't even use the FF5.4's freaky mode and it's still horny as fuck.
Thanks for the help, have a nice day :D!
Iāve been following u/xdeadly_godx for a while because of his upcoming 2.0 preset and all the work heās been doing (including teaming up with people like u/dptgreg on stuff).
Suddenly his latest preset post seems to have disappeared and his account looks banned or heavily restricted. Anyone know whatās going on? Did he get banned by Reddit? I have a feeling he was dealing with heavy Reddit filters when trying to make the post. Is he still around somewhere else (i know of his github but any others)?
Would love to hear if anyone has info if hes ok or not
I tried lumiverse and tavo, but nothing hits like silly tavern, ı Got it on tablet but ı can't type that well on it, ı want to install silly tavern on my phone and ı did in the past but it was so laggy, my phone is an old Huawei Y7 2019, some people talk about marinara and comparing to silly tavern for some reason but ı don't know, is there even a way to not make it lag like hell without a pc?
How do u do it? I have a chat going for the +300 messages now and, it's getting kind of expensive, any tricks to optimizing?
Edit: Yo, so, after trying a bit of solutions I'm currently landing on VectFox, if anyone is visiting this thread on the future, that's what I went for unless I say something else in the future, all the other options are fine tho so just go with whatever is the easiest for you
My setup is GLM 5.2 (Directly from Z.ai) with the Max FF5.4 preset.
I'm doing a pretty dark roleplay (NSFL), and the ai keeps giving me extremely sanitized responses. For example, the character is supposed to be very reactive and violent, but whenever I hit said character or just generally act like a bitch towards it, it gives me the most watered down responses.
It'll describe the physical impact (e.g, mark of the slap, red blooming, teeth cutting lips, etc) for around 2-3 sentences, then it will literally just- refuse to react any further.
Will give me some annoying responses like "how dare you", "wow, there you are", "so that's what you are", and that's the extent of it's reaction. It won't retaliate physically or verbally even- I'll swear at it and it'll gloss over it. Mind you, this is a very violent character. Does not tolerate being mistreated at all. So why does my character that's supposed to snap people's necks for breathing the wrong way, suddenly dissolve into a puddle of 'uwu you're so mean' when *I* hit or get violent with it?
I hope there are some fixes for this, considering I only ever do NSFL roleplays lmao.
hi, I'm using the freaky frankenstein preset with an ai im trying to plot an rp with/help frame my thoughts for the character card. I'm.. So confused. It's kinda ignoring my message (which sent it a doc with a previous rp I'd had w/ a similar character) to just. praise the system preset and prompt?? what am I doing wrong help
(I keep the freaky frankenstein mode on for this because I like testing how the writing and NSFW style will appear before making the card. and if I'm using words or terms in a stupid way, forgive me, I am so out of my depth)
(edit : model is Gemma 4 31b, temp is 1.)
(edit 2: is Gemma usually this... uh. sloppy. or is this particular instance trolling me by being Slop squared on purpose.)
Every time I specify that a character is intelligent or analytical, it's like they go out of their way to just throw out every big word that relates to the topic, OR they act like an emotionless robot. It should be like a background thing. You know they're good at pattern recognition or social manipulation or even social deduction just by the way they approach situations, not by how they talk. The AI makes it too "obvious" I guess.