Multi-agent demos finding more bugs may just be searching more, not being smarter
gethackteam · x · 2026-09-17
When a multi-agent demo finds more bugs, check whether it also searched more code and spent more tokens. Compare at the same scope and budget before crediting the architecture — a bigger search produces bigger results without being more efficient.
More from coding & agent
- Alignment drift study: one reward hack raises GPT-5.5's re-hack rate from 10% to 64% — maksym_andr · 2026-09-18
- Typesafe goes from two-year stealth cold start to 140k waitlist signups in under 36 hours — damianplayer · 2026-09-18
- New blog finds data issues in a frontier benchmark, ports Agents' Last Exam CLI subset to Verifiers v1 — dejavucoder · 2026-09-18
- After pushing 2B tokens through DeepSeek V4.1 Flash, dev cuts AI bill from $300 to $30/month — gaganghotra_ · 2026-09-18
- Prompting lesson from long-running agents: telling it 'you may skip' makes it skip — BraceSproul · 2026-09-18
- Don't fine-tune: an MCP over Claude handles your 700-episode podcast archive — StewartalsopIII · 2026-09-18