
Measuring token compression for coding agents on SWE-bench Lite
Compressor V2 is the combination of three independent compression strategies: brevity, tool surface reduction and tool result trimming. Each targets a different layer of an agent's request. This post measures end-to-end the gains of this new composed strategy, on real coding and tool-use workloads with paired statistical tests. The results show a combined 50% per-task cost reduction.


