Claude Opus 5 launches with Box reporting big gains on enterprise agent tasks
inductionheads · x · 2026-07-25
Anthropic’s Claude Opus 5 is presented as a major step up over Opus 4.8, and Box says it improved on complex enterprise document work in its agentic benchmark.
- Box reports gains on unstructured, end-to-end enterprise tasks such as due diligence (+17%) and life sciences target identification (+30%).
- The accompanying benchmark chart shows Opus 5 ahead on agentic terminal coding, knowledge work, search, computer use, business workflows, and biology, while also highlighting safety/alignment claims from the system card.
- The quoted release says Opus 5 is available in Claude Code/API under claude-opus-5, with the same pricing as Opus 4.8.
More from Models
- GPT-5.6 beat Claude Opus 5 on landing-page polish in a frontend test — Arindam_1729 · 2026-07-25
- Claude Opus 5 Allegedly Released with Self-Correction and Agentic Execution — mathemagic1an · 2026-07-25
- Google is called out for weak Gemma models and its stance on open-source LLMs — BoogerheadCult · 2026-07-25
- Musk says Grok 4.6 is due in 2 weeks and Grok 4.7 in 4 weeks — elonmusk · 2026-07-25
- Together says Kimi K3 matches near-flagship coding at about 35% of Claude Fable 5’s price — togethercompute · 2026-07-25
- Musk says Grok 4.5 and Claude Opus 5 are the only models on the Pareto frontier — elonmusk · 2026-07-25