New “expenditure horizon” metric compares human and agent cost efficiency on open-ended tasks

littmath · x · 2026-07-22

A proposed way to measure AI agent capability on continuously scored tasks: the expenditure horizon.

Related event: METR Proposes 'Expenditure Horizon' for AI Agent Evaluation(2 posts)→

Original post →

More from Research

Research channel →