Codex agents get stuck in endless test loops while Gemini Flash ships the same task fast

not_enough_privacy · reddit · 2026-07-27

A user says their Codex and Luna agents have become overly cautious, slow, and stuck in endless testing loops while building neurosymbolic knowledge-graph extraction systems. They describe the agents as optimizing for technical correctness and schema shape, but missing the actual content quality of the extraction.

They report getting much faster progress with Gemini Flash: it finished the same brief in about 30 minutes, after which they fixed the structural issues manually. The post asks whether others have seen the same “fearful agent” behavior and how they’ve gotten agents back to shipping useful code instead of over-testing.

Original post →

More from coding & agent

coding & agent channel →