← All guides

Manage context for AI roleplay and long conversations

Budget conversation history, character cards, and response length within a model context window.

Context includes the whole request

The model sees system instructions, character descriptions, lore, prior messages, and the new message. Your visible latest message is only one part of the input. The context cap includes input plus generated output.

Leave room for the response

For a 16,384-token context limit, a 14,000-token prompt leaves at most 2,384 tokens for output, subject to the separate output cap. Client formatting can add tokens, so leave some headroom.

Keep useful history

Trim repetitive turns or summarize older events while keeping current character and scene details. A larger context limit alone does not guarantee better memory or more consistent writing. Check model availability and limits before choosing a client preset.

Your next step

Check available models, open your console, or read the API documentation. Public signup and paid plans are still being prepared.