Claude Opus 5 tops AA-Briefcase with 1720 Elo and 20% lower task cost than Fable 5
ArtificialAnlys · x · 2026-07-25
Artificial Analysis says Claude Opus 5 is the new leader on its AA-Briefcase agentic knowledge-work benchmark, beating Claude Fable 5 by about 146 Elo at max effort.
Key results
- Max effort: 1720 Elo on AA-Briefcase, versus 1574 for Claude Fable 5
- Cost per task: $17.79, about 20% cheaper than Fable 5’s $22.30
- High-effort mode: $10.41 per task while still beating Fable 5 by 32 Elo
- Analytical quality: 2016 Elo, nearly 300 points ahead of Fable 5
- Presentation quality: 1628 Elo, still about 40 points behind GPT-5.6 Sol (max)
The benchmark is based on realistic private knowledge-work tasks with thousands of files and deliverables like reports, slides, and spreadsheets. Artificial Analysis says Opus 5’s gains come mainly from rubric pass rate and analytical quality, while the tradeoff is slower runtime and more turns per task.
More from Models
- Opus 5 gets praise for unusually strong spatial awareness — almmaasoglu · 2026-07-25
- Artist says AI-made works should be signed “AI-generated,” not hidden — Merzmensch · 2026-07-25
- Perplexity adds Opus 5 to Computer for Pro and Max users at roughly half the price — AravSrinivas · 2026-07-25
- GPT 5.6 Sol is highlighted as stronger on one benchmark row at 72.7% — himanshustwts · 2026-07-25
- Leaked Claude Opus 5 Takes 50% Longer Than Opus 4.8 in AA-Briefcase Tasks — ArtificialAnlys · 2026-07-25
- GPT-5.6 Sol looks set to beat an Act 3 A7 boss — Jsevillamol · 2026-07-25