METR Proposes 'Expenditure Horizon' for AI Agent Evaluation

METR introduced the 'expenditure horizon' metric to evaluate AI agents by comparing their cost-efficiency against humans on continuous tasks. It identifies the budget tipping point where agents outperform human cost-effectiveness.

2026-07-22 ~ 2026-07-22 · 2 related posts