Databricks says Genie Code beat three coding agents on 401 tasks at $0.55 each
matei_zaharia · x · 2026-07-26
Databricks says better data agents can improve accuracy and efficiency at the same time.
In a head-to-head test on 401 real internal data tasks, Genie Code was compared with three leading general-purpose coding agents. Each agent used its own harness, frontier models, Databricks MCP, and the same 20-minute task budget.
Reported results:
- Accuracy: 76.6%, the highest of the group
- Mean cost: $0.55 per task, the lowest of the group
- Relative cost: less than half the cost of the other agents
The company says the edge came from context handling: semantic search, persistent memory, and workspace understanding reduced random exploratory wandering and helped agents arrive at the right answer faster.
Related event: Databricks Claims Genie Code Beats Coding Agents at $0.55/Task(2 posts)→
More from coding & agent
- GPT-6 Astra beats Factorio with enemies in 44 in-game hours at ~$4,500 API cost — liminal_bardo · 2026-09-11
- 105 hidden bugs, 2 repos: DeepSeek V4.1 Flash fixes 24 at $1.80 vs Opus 5's 27 at $51.33 — ChartsJournalX · 2026-09-11
- Investment Analyst Asks How to Build a Claude-Based Diligence Agent Stack — Careless_Tie2286 · 2026-09-11
- Treating agents like 50 First Dates: a 3-layer context system so every conversation doesn't start from zero — evielync · 2026-09-11
- Running the Firefox MCP on Android via Termux, ngrok, and mcp-proxy — Nervous-Strain7544 · 2026-09-11
- SmolVM open-sources persistent computer infrastructure for agents that outlive chat sessions — aniketmaurya · 2026-09-11