UT Austin study finds token-cutting context compression can slow coding agents
A UT Austin study ran nearly 35,000 coding-agent runs on SWE-bench Verified and Terminal-Bench while varying how context is compressed, when it triggers, and how much is removed. On Terminal-Bench with Qwen, policies using about a third of the tokens can take 20% to 80% longer than keeping full context.