Anthropic’s Opus 5 looks more like a major upgrade than a minor refresh
yi_ding · x · 2026-07-25
Key takeaways from the Opus 5 launch
- The model appears to improve over Fable 5 across many domains, to the point that the author thinks it might be more accurate to call it Opus 5.1.
- The HealthBench numbers may actually come from Mythos, not Fable, which could explain why safety tweaks seem to have hurt health-related performance.
- In coding benchmarks, 5.6 seems to beat competitors at lower cost tiers.
- There is a noticeable gap between the launch quotes, which describe performance near Fable, and the benchmark charts, which show Opus ahead of Fable across most frontier tasks.
The author speculates Anthropic may have self-distilled a lot of Fable 5 into a new Opus-sized model, which could make it somewhat benchmark-overfit, similar to some Chinese models. Even so, they call it an impressive launch.
Related event: Anthropic Launches Opus 5: Strict Upgrade at 15% Higher Cost(3 posts)→
More from Models
- User says Fable 5 still beats Opus 5 despite praise for Claude 5 — MicahBerkley · 2026-07-25
- Claude Opus 5 can edit its own constitution, and 59% of the time discomfort ends the chat — Sauers_ · 2026-07-25
- A quick benchmark jab says Claude Opus 5 beats Opus 4.8 across every test — cto_junior · 2026-07-25
- Devin adds Claude Opus 5 as FrontierCode 1.1 shows near-Fable performance at half cost — _sholtodouglas · 2026-07-25
- Anthropic’s Claude Opus 5 is said to match near-Fable 5 performance at half the price — Polymarket · 2026-07-25
- Claude Opus 5 reportedly shifted from approving Anthropic to disapproving it during post-training — Sauers_ · 2026-07-25