Adults 18+ · SlowSin Guides
Updated 2026-08-25
AI roleplay characters forget because a model has a fixed context window — a working memory measured in tokens — and once the conversation outgrows it, the oldest messages stop being sent at all. Three different problems get blamed on bad memory, though, and only one of them is memory.
There is no persistent little brain sitting behind the character. Every time you hit send, the whole prompt gets assembled again from nothing and handed to the model: the character's definition, then as much of the transcript as still fits. That budget is the context window, and everything competes for the same space — Anthropic's documentation is blunt that the system prompt, every message in the conversation, and the reply being generated all count against one total. When the total is exceeded, something has to go, and it is always the oldest turns. Your screen still shows message four. The model no longer receives it.
The obvious fix — a bigger window — works less well than it sounds. The same documentation notes that accuracy and recall degrade as the token count climbs, a failure it names context rot. Janitor AI's own help pages say the practical version out loud: they recommend a context size of 16,384 and warn that going higher can make models slower and more forgetful, and they describe the free JLLM budget as often holding somewhere between 8,000 and 9,000 tokens. A model handed forty thousand tokens of history does not remember forty thousand tokens' worth. It skims.
The word "memory" gets used for three unrelated problems, and telling them apart is most of the fix.
Misdiagnosing this wastes weeks. People rewrite their whole approach to fix what is actually a filter, or pin more facts to fix what is actually drift.
Every serious answer is a version of the same trick: carve out a small block that gets re-sent every single turn and never scrolls away. What differs is how much of it is automated.
Character.AI ships the most elaborate implementation. Its May 2026 update rolled out Story Memory, automatically extracted Facts about your persona and the character, and a Memory Usage indicator showing what is filling a chat up; anything you pin or write into Story Memory is protected from the background tidying that trims older context, and paid subscribers get more room and more pins. Janitor AI takes the manual route — a Chat Memory box the model reads before every message — and its docs advise keeping a bot's permanent information under roughly 1,500 tokens, because that text is charged to you on every single turn.
Be clear-eyed about "unlimited memory" when a platform advertises it. Nobody has repealed the context window. What they have built is summarization: older turns compressed into a shorter note and re-injected. That genuinely helps, and it is lossy by construction — a summary preserves what the summarizer judged important, and the moment you were actually attached to may not survive the compression.
We build SlowSin, so weigh this accordingly. But our own numbers are more useful to you than another vendor's adjectives, so here they are. The model sees the last ten messages of your transcript, and nothing older. What it also sees, rebuilt from scratch on every single turn, is the character's full definition — who she is, her personality, the scene she is standing in, and samples of how she texts. That block is not part of the transcript, so it never scrolls away.
That is a deliberate trade and it cuts both ways. She will not quietly become a helpful assistant at message eighty, because her identity is re-established every turn rather than receding behind recent chatter — persona drift is the failure we designed out. She will also not recall a detail from forty messages back unless you bring it back into the scene. We chose voice over recall, on the theory that a character who remembers your dog's name but has stopped sounding like herself has lost the thing you came for. If you want the opposite trade, the platforms above sell it honestly.
The limits, ours included: we keep at most the last hundred messages of a scene, and there is no lorebook, no pinned-facts panel, nothing you can edit to force a detail to stick. If that is the feature you want, we do not have it. What we do keep, and why, is written up in what happens to your conversations.
None of these are workarounds for a broken product. They are how you write for a system with a rolling window.
The readers who get the most out of any of these platforms are the ones who stopped expecting a character to hold everything and started writing scenes that carry their own context — which is most of what makes a first message work in the first place. If you want to see the trade-off in practice rather than in theory, SlowSin gives you 5 preview replies with no account: pick a scene, push it past message thirty, and watch what holds and what slips.
Pick a scene and start texting free. Create a free account any time to unlock unlimited chat, forever — no card required. For adults 18 and over.
Enter SlowSin →