Databricks says Genie Code beat three coding agents on 401 real tasks at $0.55 each
DbrxMosaicAI · x · 2026-07-25
- Databricks reports that adding context can improve both accuracy and cost for data agents.
- On 401 tasks distilled from internal usage, Genie Code was benchmarked against three leading general-purpose coding agents.
- Each agent used its own harness, frontier models, Databricks MCP, and the same 20-minute budget.
- Genie Code was the most accurate agent in the test at 76.6% and also the cheapest, with a mean cost of $0.55 per task.
- Databricks attributes the advantage to semantic search, persistent memory, and workspace understanding, which reduced inefficient exploration, errors, timeouts, and cost.
More from coding & agent
- Claude Code 2.1.220 adds 6,256 prompt tokens and shifts system-token share higher — ClaudeCodeLog · 2026-07-25
- Claude Code CLI 2.1.220 ships with crash fixes and stability improvements — ClaudeCodeLog · 2026-07-25
- Grok Build update adds guided onboarding, search overrides, and failed-run resume — elonmusk · 2026-07-25
- Anthropic releases Claude Code v2.1.220 with bug fixes and reliability improvements — ashwin-ant · 2026-07-25
- A Mac coding-agent benchmark says the harness can cut token use 5× — asankhs · 2026-07-25
- Andrew Ng releases a free one-hour course on building agentic knowledge graphs — goyalshaliniuk · 2026-07-25