TraceLab: 4,300 real coding-agent sessions reveal high cache hit rates but quadratic cost growth

CShorten30 · x · 2026-08-19

A new paper TraceLab offers a large-scale characterization of real coding-agent workloads.

The quoted caching math adds: after a coffee break, a simple "hi" can cost a full dollar — your cache went cold, so you repay for the million tokens of context; even with caching, cost grows quadratically with conversation length, which is why Claude Code/Codex compact context around 200-300K tokens.

Original post →

More from coding & agent

coding & agent channel →