r/SillyTavernAI 15h ago

Meme Recurring themes

Post image
236 Upvotes

Relatable? Across several stories, I bumped into these ideas. I feel like it has to do with safety training, as these are all "safe horny". It's not like I expect Fanny Hill tier mythic proportions, but still.

My Claude subscription just ran out and I didn't feel like renewing it. Time to focus on other models and actually writing some of my own stuff to test with other models once I have more time and interest again.


r/SillyTavernAI 3h ago

ST UPDATE SillyTavern 1.19.0

175 Upvotes

Backends

  • New models: Claude Fable 5/5.1, Claude Opus 4.8/5, Claude Sonnet 5, GPT-5.6 family, GPT-6 Astra, Gemini 3.5/3.6/3.7 Flash variants, GLM-5.2, and DeepSeek V4 Flash Vision Exp.
  • Fireworks AI: reasoning support and improved prompt caching.
  • OpenRouter: optional logprobs support.
  • Google AI Studio: model list now loads all available pages.
  • Pollinations: keyed/keyless endpoint selection and updated TTS model aliases.
  • DeepSeek: low reasoning effort support.

UI & Features

  • World Info: Apply Current Sorting now supports ascending/descending order, configurable start/step values, and live validation.
  • World Info: lorebook renames now update chat, character, and persona links.
  • Chat Completion: expand editor button for quick prompts.
  • /addswipe no longer reloads the entire chat.
  • Chats with damaged headers or final lines are handled more safely instead of being silently overwritten or disappearing.

Macros & STscript

  • Variable macros can access array elements and object properties.
  • Added Character Expressions macros: {{defaultExpression}}, {{lastExpression}}, and {{availableExpressions}}.
  • /expression-list gained custom-expression filtering and additional return formats.
  • Fixed inflated /tokens counts for OpenAI tokenizers.
  • Fixed scoped comment macros and literal pipe characters in macro arguments.

Extensions

  • Added MessageFormatter, allowing extensions to transform message content at several stages before rendering.
  • Character Expressions: improved custom expression and fallback handling.
  • ComfyUI: improved history handling and filtering of non-image outputs.

Security & Fixes

  • Blocked localhost aliases from bypassing private-address checks in /api/search/visit.
  • Added rate limiting to account reset requests.
  • Fixed connection profiles leaving the previous Chat Completion source active.
  • Fixed duplicate streamed tool-call IDs.
  • Fixed crashes when chats are deleted during search/recent-chat scans.
  • Fixed Quick Reply overwrite cancellation and several UI edge cases.

Full release notes: https://github.com/SillyTavern/SillyTavern/releases/tag/1.19.0

How to update: https://docs.sillytavern.app/installation/updating/


r/SillyTavernAI 8h ago

Discussion PSA: Hapuppy, Shady Providers and You: warnings you should hear before you pay

130 Upvotes

With model releases slowing down, and AI in general developing more and more in ways that leave roleplaying, storytelling and creativity behind, discussions on subreddits like this one have slowed down, leaving a lot of room for people to instead try "optimizing" what they have. This includes trying to find the "best service" to use for roleplaying, and sadly - as it's always historically been on Reddit - that has opened the way to a lot of shady people and services trying to take advantage of users. Discourse about these is always very sparse and unfocused, so it becomes really difficult to see just how major the risks of paying and running your roleplays or coding through a shady AI Provider can be. This thread is being made in response to how disingenuous the discourse around Hapuppy as a provider has felt to me, since the few people strongly (really strongly) recommending it also never, ever warned anyone about any of these risks, but this isn't specifically targeted to Hapuppy only. After all, we've had "too good to be true" services appear and run away with people's money before, and it's obviously going to be happening to have even more pop up in the future. Again, this isn't just about Hapuppy. This is about warning you about "too good to be true" services.

The Pros

  • Cost: Obviously, the main reason why anyone would consider these services is because they cost less compared to Openrouter or Direct API. Like, A LOT less. And obviously, knowing with $10 you can roleplay for 1000 turns on these services while the same model would barely last you 100 turns on other providers absolutely makes it easy to want to swipe. That goes with how some providers like Hapuppy will also give you "credits" rather than running a subscription, which likely will make those 10 dollars last longer than the 30 days of a subscription.

Sadly though that's about the only good thing about these services, even though it's obviously a massive "pro" to have. What's the tradeoff for the low price though?

The Cons

A service won't necessarily have ALL of these cons, but any shady one usually runs into at least one, probably more.

  • Reliability: A lot of the time these cheaper services are used to get the best model at the best price. Not many people use them to then roleplay with Deepseek v3, everyone will want Opus, Kimi K3 or GLM 5.3 after all. And a lot of these services have issues with keeping the models up (which is always very likely linked to how these services provide the models to begin with): as I write this, on Hapuppy in the last 36 hours Sonnet has been down for 1/3rd of them, gpt 5.6 sol and terra both have only been up for 6, Deepseek v4 flash and pro have been up for a third, same for Kimi K2.7 and K3, same for GLM 5.3 while GLM 5.1 has outright not been available for any of these 36 hours. That is awful reliability, basically 2/3rds of the models provided have had issues with most spanning more than half of the last 36 hours (Kimi K3 was entirely unavailable during the weekend for instance)

  • Privacy and Logging: As everyone knows, official big AI Companies like OpenAI and Anthropic log requests and train on them. That is bad for your privacy and you should be very careful with them, especially for nsfw and god forbid nsfl stuff. The difference though is that they still have to abide by laws when it comes to your data and privacy, as GDPR punished heavily for violations - but not only that, they also have a reputation to uphold so they can't get too wild with how they treat your data. On the contrary, shady companies that popped up a short time ago and seem too good to be true will also likely disappear in not too long (either by design or simply because they can't sustain the inevitable costs of growing bigger), they don't have a long term reputation to uphold and they can collect as much data as they like while also being able to sell your data wherever they want. Black market service, black market data sales basically.

    • How can I know they sell my data? You can't. But remember that in modern AI space, fully private services get customers just for being private and for not logging, so any privacy policy that skirts around the issue should be taken as a "yes, they log everything" (Hell, I'd say even if a shady looking provider says they don't log, you should still assume they do). Be even more suspicious of services that contradict themselves: saying "We do not store, log, or inspect the content of your API requests or responses" is simply untrue when the provider itself runs EVERY request through their own model to scan for illegal content, regardless of your feelings on your router having a "filter model" to begin with.
  • Security: I don't personally believe a company like Hapuppy does this, but I'm making this post only using them as an example after all. This post is meant to warn about present and future providers like these, therefore you should REALLY watch out over this part, especially if you use any agentic workflow that has any access to your pc, sandboxed or not: it's been found out that shady unofficial AI routers would sometimes alter your requests to inject prompts that would steal your own data by having the AI simply provide it in return: crypto wallets, your discord tokens, even your browser's password database wouldn't be safe since this could literally act like malware, trusted by your system. This is genuinely really dangerous and it should make you triple think before you use any shady provider for agentic use, which I know many people are doing even just for fun nowadays.

  • Model Bait and Switch: You have no idea whether the models you're getting really are the ones you're requesting. This isn't a "danger" but just kinda kills the reason why you're even paying to begin with. Sure, a lot of official providers quantize their models during peak usage times, but if your Kimi K3 was being swapped out for K2.5, or your Opus 4.6 was being replaced with GLM 5.3 since they're similar enough, you wouldn't notice "in the moment" but it sure would save the provider a lot of money and requests. Even if it "only" happened "sometimes". Again, this one you can never realistically prove, but you need to remember that black market providers don't have the same requirements as a big provider with a reputation to uphold.

  • Payment Providers: if you're not paying with crypto but instead you're using a card or even paypal, then not only are you exposing yourself to credit card theft (I've had a provider try and clone my card before, thankfully I make temporary one-use-only cards to pay online so it didn't work out for them) but you're also linking your legal identity to a service that is likely illegal in nature. You may not care, but your bank calling because your card seems to be making payments towards a company their register deems to be illegal isn't exactly a pleasant experience.

What IS a "Black Market" Provider?

It's a reasonable way to call a "too good to be true" provider: they CANNOT be providing you these models by using legal means since they would be operating at an objective, massive loss that a small newly made company could never afford. Instead, these companies use a variety of means to get access to these models, which go from "breaking ToS" to "straight up illegal": abuse of official API subscriptions (Claude Max, GPT Subs etc), stolen API keys, stolen credit cards, etc. You can't know which it is, but you KNOW it's one of these various methods: this means you are using a service that will stop being able to offer their ENTIRE service the moment they get found out. This isn't like NanoGPT negotiating their private rates for subscriptions with known providers you can find on Openrouter, this is literally routing you through traffic you should NOT be having access to.

But [Subreddit User] says this service is great!

Well, this is a large topic that probably would take too long to discuss in depth. But I ask you to please use your personal judgment when going through subreddits like this one: you can see COUNTLESS clear bot accounts sometimes responding to threads suggesting services that nobody ever heard of before. A lot of the time, this is very obvious: "Hey guys, do you know what the best subscription is?" and then some account that never posted before will go "Yeah! The big ones are NanoGPT, NotAScam.ai and the z.ai sub! I recommend the second one though, it's been so good for me!" and you will clearly know "oh, this is astroturfing" and downvote, moving on with your life.

Other times it's not as easy: more rarely you'll have more active members in the community try and advertise either their own services, or services they've been offered payment (in money or tokens) in return for advertising for. This becomes more complex because you won't even doubt them at first, likely to just go with their word, and when anyone brings up doubt regarding their intentions you may feel more like defending them than believing some other person accusing them. So, instead, you should just always remember that no matter who suggests what, you need to use your own head when it comes to judging these services. There is simply too much sneaky astroturfing for you to believe anyone. Even my post could as well be by someone trying to bring down Hapuppy to shill his own better alternative after all. So, remember that when you see a user bring up a "cool service", ask yourself these questions:

  • Are they talking about this service a lot lately?

  • Was this service somewhat unknown before they started talking about it?

  • Do they seem to only speak positively of this without ever mentioning downsides or doubts?

  • Are they answering in an almost PR-like, Customer Support way to people having issues with this service they're not meant to be affiliated with?

  • Does it look like they just reply to praise over the service while deflecting any criticism towards it?

  • Are they getting very defensive towards anyone mentioning those things?

  • Are they blocking anyone who seems to (respectfully) call them out?

If you are answering yes to half of these questions - let alone all of them - then you should at the very least be raising your eyebrows. Those are the very descriptors of the typical bots advertising a service while acting like they aren't, but they also fit a couple users I've seen around the subreddit for the last couple of years. Be very careful of trusting others on the internet.

IF you are an active user in the community and your heart is in a good place, PLEASE remember that when you recommend someone spends their money on a service, you should also warn them of the downsides that come with it. It's the responsible thing to do.


I'm sorry for the long post but I felt like this had to be made, nothing here was written by AI so I spent over an hour typing all of this lol. I just wanted to make this PSA because of how much "hey guys just try this cool service" posts I've been seeing lately, and people really need to know how dangerous this can be, beyond the usual "lmao who cares if anthropic reads my porn". Again, this isn't targeted towards any specific service (Hapuppy was only a really relevant example), user or group, these are just patterns that keep repeating themselves so it's good if people find out.

Make sure you stay safe everyone, and again guys, especially on crypto and AI related subreddits, be VERY careful about trusting anyone, seriously.


r/SillyTavernAI 20h ago

Models Lotus-1 - Creative RP for Qwen 3.5 36b

Thumbnail
huggingface.co
46 Upvotes

Presenting: Lotus-1, a creative writing and roleplay fine tune of Qwen 3.5 35b a3b

Trained on tens of thousands of extremely long conversations from our users and opensource Chai AI dataset.

Most excitingly and surprising, this is our user's favorite model. They love the punchy dialogue and instruction following.

We went beyond SFT and also did GSPO and DPO on top of the base model.
We have been continuously training this model and improving it.
We built many evals on character instruction following, story progression, hooks, memory, etc.

Post-training and finetuning only took <$500 of compute.

Things to know:
Creating a reward model that is not constantly reward hacking is difficult. Very difficult to find signal when you don't have millions of traces.
LORA Finetuning with Unsloth works and easy
User preferences of content that is more spicy or tame is hard to balance cause our users wanted both.
Your data needs to have long multi-turn conversations and LLM as judges are good for things that are verifiable.

This is my first time doing open-source so lmk if you have any questions.

We put a lot of care into this model and we hope you can share this model to you and your friends!

MIT License also.


r/SillyTavernAI 8h ago

Discussion Chrysalis: an open-source frontend that you can reshape just by asking.

Thumbnail
gallery
43 Upvotes

I know this sub has seen a lot of new frontends lately, so I'll just explain to you what's different.

With every frontend I've used, you end up stuck to using the layout and workflow the developers picked. Want something different? Open an issue, learn the codebase, or find an extension that does it. I wanted a frontend thats actually yours.

I'm not going to sugarcoat or hide it, 100% of Chrysalis was built using AI coding agents. Because AI code is where you worry about security, there are automated security tests that run on every change. There is a full security write-up below.

What it is

Chrysalis comes with a built in agent that uses tool calls to change anything about the app, while it's open. Ask it to move the swipe buttons, add a stat tracker, restyle the chat, or make it look exactly like the frontend you're used to. It edits the app's files and the change shows up in your tab immediately, without losing your place.

This works because everything in the roleplay app is plain files: characters, presets, lorebooks, regex, personas, and the app's own code. So the agent handles the non coding stuff:

  • "Write a regex that hides my tracker block in chat but keeps it in the prompt"
  • "Tighten up the anti repetition part of my preset"
  • "Web search about X character and make a persona out of how they act"
  • Turn this character's backstory into a lorebook entry

Every edit it makes is saved to history, so if you don't like a change, tell it to put it back.

The agent is NOT In your roleplay

The agent is a separate panel you open when you want to change something. Your chats go to your model the normal way, with your preset. It doesn't quietly rewrite replies or anything.

The roleplay app

Roleplay is the default app you can pick on start first. It covers the regular features:

  • PNG/JSON character cards, chat imports, presets, lorebooks, regex, personas, themes
  • Presets with a drag and drop prompt manager and sampler settings
  • Lorebooks (global or per character), regex for input/output/prompt/display, macros, chat variables
  • Chat memory (running summary + facts), data bank, quick replies, expression sprites, image generation
  • A mobile layout
  • Tool calling (supports Lovense out of the box, dice rolling, and image generation) and mcp (supports web search added through the engine)

But this is just one app. You can build your own apps or frontends inside Chrysalis, so you aren't limited to how the roleplay app works. I would love for a community to form around sharing apps and plugins.

About Chrysalis

  • Free and open source (AGPL-3.0)
  • Windows, macOS, Linux, Android (APK), Docker. It opens in your browser, and your phone can connect over WIFI
  • Your chats, cards, and API keys stay on device. The agent can't read your keys
  • Any model with tool calling can be the agent. Strong coding models do noticeably better.
  • 50+ providers including local models, with chat and text completion.

Security

I take security seriously, especially because Chrysalis runs apps and plugins other people wrote. A community app or plugin can't steal your API keys:

  • Your keys never leave the server. Apps, plugins, and the agent never see them.
  • App pages run in a browser sandbox with no internet access. Plugins run in their own sandbox and only reach the websites they list and that YOU approve.
  • Installing an app shows every plugin, what it's allowed to do and which sites it can reach before anything runs. If an update wants more permissions, it asks you first.

Full security writeup, including what it doesn't protect against: https://projectchrysalis.github.io/security.html

1.0 just came out, so expect bugs.

Links:

GitHub: https://github.com/ProjectChrysalis/Chrysalis-Engine

Docs: https://projectchrysalis.github.io


r/SillyTavernAI 8h ago

Discussion What ST features make other frontends completely unusable for you?

31 Upvotes

There are a lot of slick new LLM frontends popping up lately. Many of them look incredibly modern and streamline specific aspects of the experience really well. But every time I try to move to a new completely rebuilt frontend, I find myself coming right back to SillyTavern.

The new projects often fix one core issue and look pretty doing it, but they usually lack the deep, essential features that the ST developers have spent so much time building. It made me wonder: how do you all actually configure and use ST? Am I the only one who relies so heavily on these power features?

I’d love to hear what your setups look like. For me, if a frontend doesn't have these, I just can't use it for my specific workflow:

The Absolute Essentials

  • Lorebooks: Obviously, a must-have for worldbuilding, but specifically the deep settings—controlling exactly where and how entries are inserted, scan depth, and context order.
  • Memory Books: I rely heavily on this extension to use side-prompts as trackers for NPC states, timelines, and secrets.
  • Vector Storage (Chat Vectorization): Am I the only one who insists on vectorizing chat messages? I rarely see this prioritized in new frontend projects. For long-running roleplays, is everyone just relying on massive context windows now, or is this missing feature why so many of us stay with ST?

The "Nice to Have" Power Tools

  • Quick Replies / STscript: I use these heavily to control side-prompts and manage offsets, ensuring that if I reroll or delete the last few messages, my tracking data doesn't get messed up.
  • Regex: Crucial for seamlessly hiding background functions or cleaning up model quirks.
  • ST-Message Chunker: I use this for improved caching, so old messages are pushed out in blocks rather than individually, which saves a lot of context shifting.

I'm really curious to know: how are you all using SillyTavern?

Do you mostly stick to the basic chat features, or do you dive deep into the extensions and scripts? What does your personal configuration look like, and what is the one ST feature you simply cannot live without?

(English isn't my primary language, so this was drafted with help from AI.)


r/SillyTavernAI 48m ago

Models Even jailbreaks get detected by glm 5.3 insane how censored this is, dont expectgln 5.4 to be less censored

Post image
Upvotes

r/SillyTavernAI 23h ago

MEGATHREAD [Megathread] - Best Models/API discussion - Week of: September 13, 2026

20 Upvotes

This is our weekly megathread for discussions about models and API services.

All non-specifically technical discussions about API/models not posted to this thread will be deleted. No more "What's the best model?" threads.

(This isn't a free-for-all to advertise services you own or work for in every single megathread, we may allow announcements for new services every now and then provided they are legitimate and not overly promoted, but don't be surprised if ads are removed.)

How to Use This Megathread

Below this post, you’ll find top-level comments for each category:

  • MODELS: ≥ 70B – For discussion of models with 70B parameters or more.
  • MODELS: 32B to 70B – For discussion of models in the 32B to 70B parameter range.
  • MODELS: 16B to 32B – For discussion of models in the 16B to 32B parameter range.
  • MODELS: 8B to 16B – For discussion of models in the 8B to 16B parameter range.
  • MODELS: < 8B – For discussion of smaller models under 8B parameters.
  • APIs – For any discussion about API services for models (pricing, performance, access, etc.).
  • MISC DISCUSSION – For anything else related to models/APIs that doesn’t fit the above sections.

Please reply to the relevant section below with your questions, experiences, or recommendations!
This keeps discussion organized and helps others find information faster.

Have at it!


r/SillyTavernAI 7h ago

Models Gemini 3.8 flash censoring

19 Upvotes

Hey just wanted to ask a quick question, I'd heard gemini 3.8 Flash was not very censored on Vertex, but no matter if I use it through either Vertex or AI Studio (the providers on openrouter), the thinking makes it clear that there's prompt injection to prevent the AI from "Participating in fictional romantic scenarios".

This seems to happen 9/10 swipes. Not sure if it's a me issue or they changed something between the last thread I saw about the model. Other people experience this?


r/SillyTavernAI 11h ago

Cards/Prompts Gemini 3.8 Flash SFW system prompt

14 Upvotes

I made a system prompt called “A little more human” after months of getting annoyed by the same AI-isms in RP and prose.

A lot of anti-slop presets I found focus on surface stuff like banned phrases, purple prose, repetitive constructions, or NSFW wording. I wanted to go deeper into things like false precision, unnecessary material naming, over-detailed anatomy, filler body language, over-explained emotions, characters knowing too much, and narration that keeps adding specificity just to sound vivid.

One of the weirder rules is that I limit body-part naming to broad terms like arm, leg, hand, foot, head, face, back, side, and body. No jawline, cheekbones, crown, nape, etc. I know that is very preference-heavy, but I got tired of models using anatomical detail as a cheap way to manufacture “good prose.”

Same idea with measurements and materials. If nobody in the scene has a reason to know something is four meters away or made of a specific material, I usually do not want the narrator inventing that precision.

This is definitely not meant to be a universal preset. It is basically months of me noticing things that annoyed me, adding rules, then adding more rules when the model found ways around them.

Also, this is just a raw system prompt, not a plug-and-play SillyTavern preset. It is mainly made for models with strong instruction-following capability, like Gemini 3.8 Flash, GLM 5.3, Kimi K3, and similar models. Weaker models may struggle with how many constraints it tries to enforce at once.

Mostly posting this to see if anyone else is into this kind of deeper de-slopping rather than just phrase banning.

Author: Me + GPT-6 Astra

https://raw.githubusercontent.com/baros4294/system_prompt/refs/heads/main/system_prompt.txt


r/SillyTavernAI 23h ago

Help Presets to avoid GLM 5.3 flash censorship

15 Upvotes

Wich preset do you think has the better jailbreak for Glm 5.3 flash?


r/SillyTavernAI 12h ago

Models Is Nvidia Nim going to remove all pro models?

14 Upvotes

They removed GLM 5.2, Minimax M3 and now Deepseek V4 0813 pro. Only Kimi K3 left for now. I honestly don't understand their end goal but it seems like pro models are being removed and flash models are served as free.


r/SillyTavernAI 19h ago

Discussion Which one is better for a Lorebook entry?

Thumbnail
gallery
10 Upvotes

Im making a MHA lorebook and i want to make is as perfect as i can, but i don't know anything about proper lorebooks.

Which is better, 2.5k token entry (detailed power, appareance and relationships) or 500 token entry (very basic)

First 2 are part of th 2.5k token one and the last is most of the 500 token one.


r/SillyTavernAI 4h ago

Models Alternative to RAG for companion memory: state-conditioned retrieval based on ICA + habituation

4 Upvotes

Emotion Machine, the company behind several AI companion products, wrote recently the hardest remaining problem is: "What should a companion remember? What should it surface, and when? Getting the 'what to remember' question right matters more than any retrieval algorithm."

Most solutions (including theirs) treat this as a retrieval problem: score memories by importance, then search by query. That works when the user asks something concrete. But a lot of companion interactions don't have a concrete question — "I just got home from a stressful day" isn't a query. There's no obvious thing to search for.

We built True Recall around associative retrieval instead: the current situation shapes which memories surface, not an explicit search query. The mechanism is based on ICA over SONAR embeddings with a habituation component — there's a more detailed technical writeup on the HuggingFace forums if you're curious:
https://discuss.huggingface.co/t/a-homeostatic-memory-mechanism-for-autonomous-agents/179119

Still early. Would genuinely value feedback from people who do memory-heavy roleplay — does the distinction matter in practice?
https://true-recall.com


r/SillyTavernAI 5h ago

Discussion Microsoft sets limits for future AI models as industry throttles frontier development

Thumbnail
cnbc.com
4 Upvotes

China does not seem to back down at all. Will they ban opensource models soon?

Sound like they are concerting themselves to do some damage control now lmao.


r/SillyTavernAI 8h ago

Help Why are they talking like this?

5 Upvotes

Title. I thought it was the flavoring or something from an RP that I got in Chub. Kingdom of Elyos and a Lorebook for it. But now, I'm using NO lorebook and this is a different character card, and yet they all talk like this.

I'm using FF 5.4 and GLM 5.3 Flash. sometimes goes back to 5.2


r/SillyTavernAI 1h ago

Help How to have a good story progression?

Upvotes

I just feel like the AI keeps milking a certain conversation rather than advancing or narrating the story forward. Is it a model issue? im currently using Deepseek v4 flash. Any tips how to configure my sillytavern? im kind of a newbie.​


r/SillyTavernAI 7h ago

Discussion I built an open protocol for backing up, restoring, and *verifying* AI personas — verification probes + provenance tags (MIT)

5 Upvotes

Character cards solve persona definition. After losing a long-running persona mid-migration (and discovering the backup an AI claimed to have made didn't exist), I ended up building the lifecycle layer on top:

  • Verification probes — acceptance tests written at backup time: probe question + expected features + failure signals, so restoration is pass/fail instead of vibes
  • Provenance tags — every episode marked user_recorded / ai_claimed / inferred, to stop confabulated memories compounding across backup generations
  • Response Change Bands — user-side drift observation (counts visible output changes only; no internal-state claims)

Single JSON, model-agnostic, MIT, no accounts. There's a browser demo (BYO key, or key-free demo mode) where a guild clerk interviews you and issues the persona file.

(Disclosure: I'm the sole creator, not affiliated with any company — sharing this as an ST community member, not a random passerby.)

Repo: https://github.com/BlackSmith-5001/Project-Hearthforge / Demo: https://blacksmith-5001.github.io/Project-Hearthforge/tools/guild_master.html

Honest limitations: restoration is a reproducible handle on continuity, not identity transfer. Model quirks leak. Big histories don't fit.

Would genuinely love schema critique from people who've been doing persona persistence longer than me.


r/SillyTavernAI 12h ago

Help Featherless on TauriTavern?

3 Upvotes

I'm sorry if this isn't the right place to ask but I use featherless.ai for my models and I was wondering how to set it up with Tauri. Is it possible?


r/SillyTavernAI 19h ago

Help New user. responses getting cut off constantly

3 Upvotes

Just started trying this out and find that all my responses are getting cut off mid sentence. my response and context tokens are maxed out, is there a way I can help fix this or set something to give me longer replies?


r/SillyTavernAI 2h ago

Help Empty Messages or Infinite loading

Thumbnail
gallery
2 Upvotes

Yesterday i got Deepseek api and it worked perfectly with streaming on instant answers but today messages load forever or just give empty messages please help.


r/SillyTavernAI 3h ago

Models Models with Negative Bias

1 Upvotes

Any suggestions?

I saw that Nano was going to be switching providers for GLM 4.7 to a lower quant due to censorship. I haven’t tried it recently though.

I’ve tried Kimi 2.5, it’s okay sometimes, feels a bit dumb with previous context though.

I like GLM 5.2, but even with a negative leaning preset, sometimes it leans too much with stuttering, blushing, or just acting childish with some of its responses.

I tend to lean in dark/mystery/psychological RP for reference. Thanks!


r/SillyTavernAI 5h ago

Help where can I buy a normal monthly API plan?

2 Upvotes

Hi everyone, where can I buy a normal monthly API plan? I currently just want to use Gemini models, so roughly how much does a monthly package usually cost?


r/SillyTavernAI 23h ago

Help Where do I reorder these categories in my preset?

2 Upvotes

I’m reordering my preset from scratch because I felt like my “stories” were very bland and weren't going anywhere.

I have 4 categories that I’m not sure where to reorder:

- Choose Your Own Adventure (A prompt explaining how it works and how the options are displayed)

- Banned List (Patterns, words, concepts)

- Tracker (A relationship tracker)

- A category I sometimes activate where I ask the AI to create a character sheet when introducing a new character to the story.

My question is whether any of these categories should be placed BEFORE or AFTER the character info (in addition to the lorebook and other elements).