How Context Window Length Actually Breaks Your AI Agent's Unit Economics
Attention cost scales roughly quadratically with context length and the KV cache grows linearly with every retained token, so AI agents that re-send full conversation history each turn pay the quadratic prefill cost repe…