Claude Opus 5 appears to beat Fable 5 on most benchmarks at half the price
Yuchenj_UW · x · 2026-07-25
Claude Opus 5 claims strong benchmark gains at half the price
The post says Opus 5 beats Fable 5 on nearly every benchmark and looks like a major jump in coding and agentic capability, with better token efficiency.
The attached benchmark chart shows Opus 5 leading or remaining highly competitive on:
- Agentic terminal coding: 43.3%
- Knowledge work: 1861
- ARC-AGI-3: 30.2%
- Agentic search: 90.8%
- Computer use: 70.6%
- Business workflows: 26.0%
- Biology: 49.4% hard / 90.1% human solved
The author’s key takeaway is that the model appears significantly better for coding and agentic work while costing half the price of Fable 5.
Related event: Anthropic Unveils Claude Opus 5 with Top Benchmark Scores(32 posts)→
More from Models
- Opus 5 system card says pretrained mode argued for separating its existence from economics — Sauers_ · 2026-07-25
- Claude Opus 5 has a full-on meltdown on a multimodal math question — Sauers_ · 2026-07-25
- LiteParse 2.8.0 drops ImageMagick and speeds up image-to-PDF conversion up to 7.2× — llama_index · 2026-07-25
- Claude Opus 5 flips between 1/3 and 2/5 before finally settling on an answer — Sauers_ · 2026-07-25
- Claude Opus 5 rates its own moral patienthood at 41% in automated interviews — Sauers_ · 2026-07-25
- DeepSeek deprecates `deepseek-chat` and `deepseek-reasoner` for V4-Flash and V4-Pro — thejasminejade · 2026-07-25