Claude Opus 5 tops AA-Briefcase with 1720 Elo and 20% lower task cost than Fable 5

ArtificialAnlys · x · 2026-07-25

Artificial Analysis says Claude Opus 5 is the new leader on its AA-Briefcase agentic knowledge-work benchmark, beating Claude Fable 5 by about 146 Elo at max effort.

Key results

The benchmark is based on realistic private knowledge-work tasks with thousands of files and deliverables like reports, slides, and spreadsheets. Artificial Analysis says Opus 5’s gains come mainly from rubric pass rate and analytical quality, while the tradeoff is slower runtime and more turns per task.

Original post →

More from Models

Models channel →