LLMs estimate human task time 3-4x longer than their own, study finds

maksym_andr · x · 2026-08-18

Research indicates LLMs are developing time awareness. When estimating task completion time for a "human expert," models like Claude Code and Codex predict durations 3 to 4 times longer than their own execution time. The study evaluates agents on long-horizon benchmarks (ProgramBench, PaperBench) to assess their ability to predict wall-clock time.

Related event: New Oxford Research Finds LLM Agents Largely Lack Time Awareness(9 posts)→

Original post →

More from Models

Models channel →