Databricks rolled Astra out to all 3,500 engineers: beats Opus 5 on complex tasks, +60% coding spend
gdb · x · 2026-09-17
Databricks exec Peter Wendell shared first-hand notes after deploying Astra to all 3,500 engineers, amplified by gdb.
- On highly complex tasks (high-level system design, long-range horizontal work), Astra unambiguously outperforms their previous top models, Opus 5 and Sol 5.6.
- Engineers given Astra increased overall coding spend by 60% vs baseline.
- No meaningful gains on medium/low-complexity tasks — they suspect those are already saturated by earlier models.
- Findings came from piloting with 200 users to gather quality and cost signal.
A rare large-scale enterprise data point: frontier model value is concentrating on the hardest long-horizon tasks.
More from coding & agent
- Agent services: a monitorable sandbox-escape relay whose logs unlock after one week — cis_female · 2026-09-17
- SciML lead stacks Claude, Codex, Devin, Grok subscriptions to maintain libraries—and it's still not enough — ChrisRackauckas · 2026-09-17
- Quick Start: Building Your Own Claude Skills With Just a SKILL.md — technextpreneur · 2026-09-17
- Atlassian design lead built a Swift personal workbench to orchestrate dozens of local AI agents — davidhoang · 2026-09-17
- Nat Friedman's Muse impresses with sub-150ms responses so fast users suspect a glitch — manosaie · 2026-09-17
- Dev Vibecodes Procedural First-Person Hands, Releases Free Game Mod — TAbrodi · 2026-09-17