Benchmark chart shows Claude Opus 5 ahead on coding, search, and biology tasks
legit_api · x · 2026-07-25
A repost highlights a benchmark chart for Claude Opus 5, showing the model stacked against Fable 5, Opus 4.8, and GPT-5.6 Sol across agentic coding, knowledge work, search, computer use, workflows, health, and biology.
The image suggests Opus 5 is a significant upgrade, especially on agentic coding and several reasoning-style tasks, though some task rows still show close competition with other frontier models.
Related event: Anthropic Unveils Claude Opus 5 with Top Benchmark Scores(28 posts)→
More from Models
- Vercel adds Claude Opus 5 to AI Gateway with fast mode for coding agents — EricBuess · 2026-07-25
- Early Claude Opus 5 feedback says it helps ship PRs faster in Claude Code — EricBuess · 2026-07-25
- Claude Opus 5 feels like Fable, but cheaper — cedric_chee · 2026-07-25
- Anthropic says Opus 5 still trails Mythos 5 on exploit generation despite better vulnerability finding — TheZvi · 2026-07-25
- Anthropic’s Opus 5 looks more like a major upgrade than a minor refresh — yi_ding · 2026-07-25
- Anthropic launches Claude Opus 5, with blind tests placing it above GPT-5.6 — lennysan · 2026-07-25