3 posts
The art of giving context: how much information should you give an AI model? Too little context yields incomplete answers, too much creates noise. A guide to relevant selection, ordering, and measurement.
What is a context window? A context window is the maximum length of text, measured in tokens, that a language model can process at once and take into account while generating a response. This guide: a clear definition, how it works, token limit, long context, memory management, the need for RAG, model comparison, and FAQs.
What is a token? A token is the smallest unit of meaning a language model uses to process text — it can be a word, a word piece, or punctuation. This guide: a clear definition, how tokenization works, the token–context window relationship, LLM cost, and why API pricing is token-based.