Anthropic says Claude Opus 5 is now state of the art on coding and knowledge-work evals
claudeai · x · 2026-07-25
In a reply, Anthropic says Claude Opus 5 is the new state of the art on several coding and knowledge-work evaluations.
This post adds a performance framing to the launch: beyond the product announcement, Anthropic is explicitly claiming benchmark leadership in practical work tasks.
Related event: Anthropic Releases Claude Opus 5(40 posts)→
More from Models
- Polymarket puts U.S. AI safety bill odds at 34% as Claude Opus 5 surfaces — Polymarket · 2026-07-25
- Moonshot AI launches Kimi K3 with 2.8T parameters and a 1M-token context — dl_weekly · 2026-07-25
- Anthropic’s Claude releases appear to have sped up from every four months to monthly in 2026 — dustinvtran · 2026-07-25
- NVIDIA reveals the winners of its Nemotron Reasoning Challenge — NVIDIAAI · 2026-07-25
- Claude Opus 5 lands on AWS Bedrock with ZDR and production APIs — AWS ML Blog · 2026-07-25
- Opus 5 is shown as a new Pareto-optimal LLM with strong ARC-AGI-3 results — brandon_galang · 2026-07-25