Claude says Opus 5 is now the new SOTA on coding and knowledge-work benchmarks
chu_onthis · x · 2026-07-25
A repost citing Claude AI says Opus 5 is now the new state of the art on several coding and knowledge-work evaluations. The post itself adds only a brief “we cooked again” reaction, but the substantive claim is the benchmark result.
Related event: Anthropic Launches Claude Opus 5 with Leading Benchmark Performance(75 posts)→
More from Models
- Early tests say Claude Opus fumbles content and strategy, while Grok 4.5 wins — JOBhakdi · 2026-07-25
- Anthropic meme says Opus 3 survived because newer defaults are even worse — repligate · 2026-07-25
- Claude Opus 5 Reportedly Falls Back to Opus 4.8 for Cybersecurity Requests — rez0__ · 2026-07-25
- GPT-5.6 Sol edges Opus 5 on DeepSWE with 72.7% vs 68.8% — rohanpaul_ai · 2026-07-25
- Claude Opus 5 lands on Google Cloud Agent Platform with $100 monthly credits — rseroter · 2026-07-25
- OpenRouter adds xAI’s Grok STT with 25 languages and $0.10/hour pricing — SpaceXAI · 2026-07-25