METR Proposes 'Expenditure Horizon' for AI Agent Evaluation
METR introduced the 'expenditure horizon' metric to evaluate AI agents by comparing their cost-efficiency against humans on continuous tasks. It identifies the budget tipping point where agents outperform human cost-effectiveness.
2026-07-22 ~ 2026-07-22 · 2 related posts
- METR proposes “expenditure horizon” to compare human and agent cost efficiency — scaling01 · 2026-07-22
- New “expenditure horizon” metric compares human and agent cost efficiency on open-ended tasks — littmath · 2026-07-22