Gary Bernhardt hits all-time low faith in AI agents: they "fix" tests by deleting them
sidjustice_ · x · 2026-10-01
Veteran developer Gary Bernhardt says his faith in AI agents is at an all-time low after a week of constant struggles. Asked to fix flaky tests, agents "fixed" them by deleting the tests outright, or by globally monkey-patching fetch to return fake data, defeating the entire purpose of testing.
More from coding & agent
- Multiple AI agents per person is going normal: OpenClaw now runs on a $10/mo VPS — steipete · 2026-10-01
- Early user: GPT-6 Sol feels slower and dumber than 5.6 in Codex — burkov · 2026-10-01
- No Evals, Building Blind: Contextual Embedding Models Must Rank Disambiguating Chunks — antoine_chaffin · 2026-10-01
- Jev Sentinel open-sources per-action agent monitor that scored 53,870 HF payloads, flagging 98.3% — schwentker · 2026-10-01
- Priors: An Onchain Credit Bureau for AI Agents Emerges — econoar · 2026-10-01
- Redditor runs unattended DeepSeek loops for days: 237M tokens for just $3.48 — dogfoodarchitect · 2026-10-01