Isolated review context saves 33% tokens
Putting a search agent's final review in a separate context keeps comparable performance across four tests while using 33% fewer tokens, the authors claim.
Put a search agent's final review in a separate context and it uses 33% fewer tokens while keeping comparable performance across four tests.
Previously search execution and final review were often carried by the same context, so the model was easily swayed by its own earlier decisions and the review could not be independent.
The authors claim a model that took part struggles to judge its own prior decisions objectively; with isolated contexts performance across the four tests was comparable and tokens fell 33%, measured by the authors themselves.
The result has not yet been reproduced or compared item by item by other teams, and cross-model reproduction and component-ablation experiments are still needed; it reports only an agent phenomenon and cannot be directly extrapolated to the same mechanism in human investors. The paper was submitted on 24 August, is marked EMNLP 2026, and its version and experimental claims are at arXiv:2608.23045.