Treating context as a budget you spend is the right mental model. Every token you burn on stale history is one the model cannot use for the actual task, and past a point more context makes answers worse, not better. I now prune aggressively and keep only what the current step needs.