capy beats a major coding harness at half the cost and half the time, eval finds
garrytan · x · 2026-09-22
A developer ran an eval comparing capy against a major coding harness: capy performed better at half the cost and half the time. This follows Garry Tan's praise that capy handles multi-step workflows and large PRs faster than Codex or Claude Code.
More from coding & agent
- Allie Miller: companies overinvest in AI productivity, ignore workflow handoffs — alliekmiller · 2026-09-22
- OpenAI's Logan Kilpatrick: AI product teams should spend >25% of time on benchmarks — OfficialLoganK · 2026-09-22
- Dev swaps in-game 3D models with Scenario's MCP right from his harness — AIandDesign · 2026-09-22
- Cua AI releases Cua-Bench-S1 benchmark and Cua-S1-Nano/4B computer-use models — ycombinator · 2026-09-22
- Multi-agent scaling may dominate next, calls for OpenAI to publish curves — 1a3orn · 2026-09-22
- GPT-6 prompt cache survives reasoning-effort changes; Codex adapts effort mid-CoT — daniel_mac8 · 2026-09-22