Loading…
Loading…
Written by Max Zeshut
Founder at Agentmelt · Last updated Sep 9, 2026
The maximum number of tokens a language model can process in a single call—encompassing the system prompt, conversation history, retrieved documents, tool outputs, and the model's response. Context windows range from 8K tokens (older models) to 200K+ tokens (Claude, Gemini). A larger context window allows agents to reason over more information simultaneously, but cost scales linearly with input tokens. Understanding your model's context window is essential for designing retrieval strategies and conversation management.
See it as a workflow
Contract Review WorkflowTrigger, steps, n8n nodes, guardrails and an importable template — plus what it costs to have it built.
Or skip the build
Workflows from $197/month, custom agents from $2,000.