Epoch Benchmarks: GPT-5.6 Latency Grows Quadratically, Claude Near-Linear at Long Context

Epoch AI measured first-token latency up to 1M tokens: GPT-5.6's latency grows quadratically with context while Claude models remain near-linear, a key consideration for long-running agents.

2026-09-16 ~ 2026-09-17 · 2 related posts