Databricks says Claude Opus 5 tops its coding benchmark and ships in AI Gateway
thesaraharminta · x · 2026-07-25
Databricks says Claude Opus 5 is now available in its AI Gateway, with governed access to data already in the Lakehouse.
An internal Databricks evaluation claims Opus 5 is a clear step up from Opus 4.8:
- Coding: top score on their internal coding bench, beating even Fable 5 while costing about the same as Opus 4.8.
- OfficeQA Pro: 60.9 vs 48.12 for Opus 4.8, and roughly on par with Fable 5.
- Enterprise agentic search: ranked #4 on the bench, but still clearly ahead of Opus 4.8.
Databricks says the model brings better agentic coding, knowledge-work performance, and long-horizon reasoning, with stronger performance per token and lower cost at production scale.
More from Companies & People
- IIT Madras Launches EdTech Tulna Standards for AI-Powered Learning Products — ravi_iitm · 2026-09-11
- Warp's six non-engineering teams all run on Linear and Claude Code — mon__lim · 2026-09-11
- Anthropic Insiders: Not Everyone at the Lab Believes in High p(doom) — anpaure · 2026-09-11
- Lumara AI Film Festival Comes to NYC Oct 26, Top AI Filmmakers to Compete — 0xAllen_ · 2026-09-11
- X drama: Anthropic researchers accused of spying on academic customers and racing them to results — basedjensen · 2026-09-11
- Investor argues Palantir-Nvidia partnership should slash Anthropic's IPO valuation — pdamodaran · 2026-09-11