Manage context for AI roleplay and long conversations
Budget conversation history, character cards, and response length within a model context window.
Context includes the whole request
The model sees system instructions, character descriptions, lore, prior messages, and the new message. Your visible latest message is only one part of the input. The context cap includes input plus generated output.
Leave room for the response
For a 16,384-token context limit, a 14,000-token prompt leaves at most 2,384 tokens for output, subject to the separate output cap. Client formatting can add tokens, so leave some headroom.
Keep useful history
Trim repetitive turns or summarize older events while keeping current character and scene details. A larger context limit alone does not guarantee better memory or more consistent writing. Check model availability and limits before choosing a client preset.
Your next step
Check available models, open your console, or read the API documentation. Public signup and paid plans are still being prepared.