[-] MindlessMass@lemmy.world 1 points 3 days ago

Probably trolls, I've had to start sorting posts by active instead of top.

[-] MindlessMass@lemmy.world 2 points 3 days ago

if you click the edit button in the top-right corner, you can ask the AI helper questions like that and it will examine the code of the generator for you to try and give you answers. I have been using it to update/modify/add features to ai-character-chat and it has been helping a ton. It does get confused sometimes but it has been really helpful for understanding how things work and/or are structured in the code and the way the AI handles some stuff. Here is an answer from the AI helper on perchance.org/ai-rpg: `Short answer: no, not in the way you're imagining — and here's why.

The "memory" that runs out is not server storage. It's the AI's context window: a fixed budget of tokens (roughly "pieces of text") the model can look at on each reply. Anything you want the AI to remember has to be placed inside that window every single time it writes. So a long recap file on your PC would only help the AI if its contents were loaded into the prompt — and the moment they are, they consume exactly the same memory as if the text lived on a server. Relocating the text doesn't free up memory; only shrinking it does.

That's the practical fix, and it's simpler than it sounds. The goal is to spend fewer tokens on your recap, not to store it elsewhere:

Compress, don't collect. Replace the long "review" with a tight bullet-point summary of only what truly matters (named characters, their goals, promises, secrets, unresolved plot threads). Dense facts > prose. The same adventure can usually fit in a quarter of the space.
Periodically summarize the summary. Every so often, ask the generator itself: "Summarize the story so far in 100 words, keeping only the most important facts." Then replace the old recap with that shorter version. This is exactly how the technique works in professional LLM apps — it's called summarization/compression, and it's the legitimate workaround for a finite context window.
Trim history. If the generator feeds the whole chat log into the prompt, the old messages are the real memory-eater. Shortening or removing early messages (after condensing them into your recap) recovers memory far more effectively than any file trick.

So the accurate mental model: your recap isn't "stored memory," it's a payment you make from a fixed token budget on every message. You can't move the budget — but you can spend it much more efficiently.` Good luck

MindlessMass

joined 2 years ago