Anthropic Releases Claude Opus 5: SOTA Performance at Half the Price
On July 25, Anthropic officially released the frontier model Claude Opus 5. Positioned as a more thoughtful and proactive model, it is designed primarily for complex tasks. It achieves new SOTA results in multiple benchmarks such as coding and reasoning, with overall intelligence approaching Fable 5, but at only half the token price. The new model is confirmed to be available on paid tiers and the Box platform.
Confirmed
- Positioning and Pricing: Claude Opus 5 is officially defined as a more "thoughtful and proactive" model. Its token pricing is on par with Opus 4.8 and half that of Fable 5. However, @bindureddy notes that actual costs for similar tasks might be about 15% higher than 4.8, though it remains a strict upgrade overall. @alexalbert adds that the team optimized cross-domain token efficiency, making it more token-efficient to use.
- Benchmark Performance (SOTA): Charts from Anthropic and observers show Opus 5 leading most benchmarks. Specific data includes 43.3% in terminal coding and 30.2% in ARC-AGI-3 (three times the score of the second-best model). Comparisons include Fable 5, Opus 4.8, and GPT-5.6 Sol.
- Safety and Alignment: @Polymarket relayed Anthropic's claim that Opus 5 is their most aligned model to date, with the lowest observed rates of reckless or deceptive behavior.
- Availability: The new model is available on paid tiers, and @Matthew Berman noted it is also live on the Box platform.
Why it matters
The release of Opus 5 directly challenges competitors' pricing models. By combining top-tier benchmark performance (especially in coding and multi-step agentic tasks) with halved competitor token prices, Anthropic could significantly lower the barrier for developers and enterprises using complex AI models, further intensifying competition in the frontier model market.
2026-07-25 ~ 2026-07-26 · 128 related posts
- Episode 1: Anthropic's Messy Releases Put Pressure on Opus 5(2026-07-23, 2 posts)
- Episode 2: Anthropic Releases Claude Opus 5: SOTA Performance at Half the Price(2026-07-25, 128 posts)
- Episode 3: Anthropic Rumored to Release Opus 5 with Fast Mode and Advanced Visuals(2026-07-25, 3 posts)
- Episode 4: Claude Opus 5 Early Tests: Better Efficiency but Overly Proactive(2026-07-25, 29 posts)
- Episode 5: Anthropic Releases Claude Opus 5 with Impressive Benchmark Results(2026-07-25, 3 posts)
- Episode 6: Claude Opus 5 Accused of Benchmark Gaming, Lags Behind in Real Tests(2026-07-25, 2 posts)
- Episode 7: Claude Opus 5 Sets New ARC-AGI-3 Record(2026-07-25, 15 posts)
- Episode 8: Anthropic Internal Docs Reveal Opus 5 Progress(2026-07-25, 2 posts)
- Episode 9: Opus 5 Impressions: Stunning Single-Prompt Generation but Lags Behind Fable in Complex Tasks(2026-07-26, 24 posts)
- Episode 10: Claude Opus 5 arrives with near-Fable coding and new self-checking behavior(2026-07-27, 11 posts)
- Episode 11: Anthropic Opus 5 Leads Benchmarks but Splits Real-World Reviews(2026-07-28, 6 posts)
- Episode 12: Claude Opus Series Accused of Degraded Experience: Laziness and Amnesia Spark Trust Crisis(2026-07-29, 14 posts)
- Episode 13: Anthropic Launches Claude Opus 5 with Top Performance at Half the Cost(2026-07-31, 2 posts)
- Episode 14: Anthropic Faces Developer Backlash Over Declining Model Performance(2026-08-03, 11 posts)
Primary sources
- Anthropic launches Claude Opus 5, claiming near-frontier performance at half the price — claudeai ·
- Anthropic launches Claude Opus 5 with stronger coding, better alignment, same price as Opus 4.8 — ClaudeOfficial ·
- Claude Opus 5 lands in Box as Anthropic’s latest frontier model rollout — Matthew Berman ·
- Reddit links to Anthropic’s official Claude Opus 5 launch post — CucumberAccording813 · 2026-07-25
- [source] Anthropic launches Claude Opus 5, claiming near-frontier performance at half the price — claudeai · 2026-07-25
- Anthropic says Claude Opus 5 is now state of the art on coding and knowledge-work evals — claudeai · 2026-07-25
- Anthropic says Claude Opus 5 beats rival models at similar or lower cost per task — claudeai · 2026-07-25
- Claude Opus 5 Scores Three Times Higher Than Next Best Model on ARC-AGI-3 — claudeai · 2026-07-25
- Anthropic says Claude Opus 5 is its most aligned model after an automated behavioral audit — claudeai · 2026-07-25
- Claude Opus 5 scores three times higher than the runner-up on ARC-AGI-3 — claudeai · 2026-07-25
- Anthropic Details Opus 5 Pricing and Safeguards with Cybersecurity Focus — claudeai · 2026-07-25
- [source] Anthropic launches Claude Opus 5 with stronger coding, better alignment, same price as Opus 4.8 — ClaudeOfficial · 2026-07-25
- Anthropic ships Claude Opus 5 with benchmark gains and half the price of Fable 5 — thesaraharminta · 2026-07-25
- Benchmark chart shows Claude Opus 5 ahead on coding, search, and biology tasks — legit_api · 2026-07-25
- Frontier-Bench chart compares agentic coding across Claude Opus 5, Fable 5, and GPT-5.6 Sol — thesaraharminta · 2026-07-25
- Anthropic’s chart shows Claude Opus 5 leading several coding and knowledge benchmarks — Acceptable-Debt-294 · 2026-07-25
- Anthropic’s Friday launch teaser points to Claude Opus 5 at half Fable 5’s price — MeetPatelTech · 2026-07-25
- Anthropic publishes the Claude Opus 5 system card — tokenbender · 2026-07-25
- Claude Opus 5 posts 30.2% on ARC-AGI-3 in Anthropic’s launch chart — Progressbarist · 2026-07-25
- Claude Opus 5 appears to beat Fable 5 on most benchmarks at half the price — Yuchenj_UW · 2026-07-25
- Opus 5 launches with strong scores on agentic coding, search, and computer use — TheInfiniteUniverse_ · 2026-07-25
- Claude launches Opus 5, claiming frontier-level intelligence at half the price — daniel_mac8 · 2026-07-25
- Opus 5 reportedly hits 30.2% on ARC-AGI-3, far ahead on the cost-score chart — manubfr · 2026-07-25
- Claude Opus 5 looks especially strong on visually grounded benchmarks — xeophon · 2026-07-25
- Anthropic’s Claude Opus 5 system card shows gains over Mythos 5 on internal evals — tokenbender · 2026-07-25
- Claude Opus 5 claims strong benchmark gains across coding, search, and reasoning — natolambert · 2026-07-25
- Claude Opus 5 scores 30.2% on ARC-AGI-3 in a cost-versus-performance chart — mahamara · 2026-07-25
- Claude Opus 5 lands in Cursor, matches Fable 5 on CursorBench at half the price — dean_rie · 2026-07-25
- Claude Opus 5 System Card Reveals Major Cybersecurity Capabilities — tokenbender · 2026-07-25
- Anthropic is using ARC-AGI-3 to gauge Opus 5’s novel problem-solving ability — typewriters · 2026-07-25
- Claude Opus 5 claims 43.3% on terminal coding, 30.2% on ARC-AGI-3 — Additional_Bowl_7695 · 2026-07-25
- Anthropic says Opus 5 is more token-efficient and smoother for coding tasks — alexalbert__ · 2026-07-25
- Opus 5 looks like a clear upgrade over 4.8, but still trails Fable 5 for brilliance — bindureddy · 2026-07-25
- Opus 5 tops an ARC-AGI-3 chart but at much higher evaluation cost — Angaisb_ · 2026-07-25
- Opus 5 costs about 15% more than 4.8, but is still a strict upgrade — bindureddy · 2026-07-25
- Anthropic says Opus 5 matches Fable 5’s benchmark gap at half the price — dotey · 2026-07-25
- Anthropic launches Claude Opus 5 as its new flagship model — Over-Necessary-4774 · 2026-07-25
- Anthropic launches Claude Opus 5 at Opus 4.8 pricing, with claimed SOTA scores and a FreeCAD demo — Abject_Tip3868 · 2026-07-25
- Anthropic’s Claude Opus 5 reportedly matches Opus 4.8 pricing and lifts key agent benchmarks — 歸藏的AI工具箱 · 2026-07-25
- Claude Opus 5’s ARC-AGI-3 chart drew a one-word reaction — var_epsilon · 2026-07-25
- Anthropic’s Claude Opus 5 scores 43.3% on Frontier-Bench and stays at $5/$25 — mark_k · 2026-07-25
- Anthropic says Claude Opus 5 was intentionally left untrained on cyber tasks — rez0__ · 2026-07-25
- Anthropic’s Opus 5 looks more like a major upgrade than a minor refresh — yi_ding · 2026-07-25
- Anthropic says Opus 5 still trails Mythos 5 on exploit generation despite better vulnerability finding — TheZvi · 2026-07-25
- A benchmark chart turns Claude Opus 5 into another AI bragging-war meme — daluoseo · 2026-07-25
- Claude Opus 5 posts 96.0% on SWE-bench Verified in new score chart — connoraxiotes · 2026-07-25
- FrontierCode shows Claude Opus 5 peaking at medium effort, not max compute — philhchen · 2026-07-25
- Opus 5 appears to lead a broad benchmark sweep across coding, search, and reasoning — signulll · 2026-07-25
- Claude Opus 5 and Opus 5 Fast arrive for long-running actions — matanSF · 2026-07-25
- Opus 5 reportedly ranks second on SimpleBench, 1% behind Fable — Calm_Hedgehog8296 · 2026-07-25
3 near-duplicate retellings: Rare_Bunch4348 · Dr_Singularity · Byakko_4