Cheap model's overnight bug-fix PRs: only 6 of 14 survived a stronger model's review
Aggressive-Narwhal-3 · reddit · 2026-10-03
An experiment on the author's own repos: during an API discount window, a cheap DeepSeek model ground through audit-found bugs overnight.
- Round 1: 14 PRs, all re-reviewed by GPT-6.1 Sol; 6 merged after CI, 8 closed. Failures included a fix that broke other valid inputs, tests needing tools absent from CI, a bypass that still worked, a key test never run in CI, and a new test that couldn't catch the bug even when code was deliberately broken.
- Round 2 flipped roles: a stronger model designs each fix, another reviews the design, and the cheap model only implements. So far 7 PRs, none approved yet.
- The author asks the community where cheap-model fixes fail and what plan/implement/review split actually works.
More from coding & agent
- Lucy hits 94.7% Recall@10 on FinanceBench with Qdrant over 2.8M SEC/DART filings — qdrant_engine · 2026-10-03
- The AI slop loop: agents rewriting each other's junk, and thin context is the root cause — IgorCarron · 2026-10-03
- 8 core networking concepts every developer should understand — goyalshaliniuk · 2026-10-03
- Codex vs Claude Opus head-to-head: porting retro console games from CPU to GPU rendering — ssh4net · 2026-10-03
- Open-source iFixAi audits AI agents in 120s with 60 checks and an A-F grade — thisdudelikesAI · 2026-10-03
- Git doesn't fit parallel agents: worktrees eat disk, so agents will design the alternative — Al_Grigor · 2026-10-03