Elo-per-token Analysis Explains Why LLM Agents Scale Fast Then Slow Down

Kaiyuan Liu · hf · 2026-09-15

This paper introduces Elo-per-token analysis to study LLM agents' test-time strategies. Agents initially scale faster than independent sampling but eventually slow down; parallel short sessions outperform a single long run.

Original post →

More from Research

Research channel →