Long-running agents: error compounding, context pollution and weak self-correction

Sad_Lavishness_53 · reddit · 2026-10-05

A structured Reddit discussion on why agents collapse on long tasks even when every single step is easy:

The core open question: is the fix better models, or better scaffolding — checkpoints, verifier steps, state stored outside the context window? The author asks for practitioner input on history pruning vs. full context, whether separate critic/verifier models actually pay off, and where agents typically start breaking in real tasks.

Original post →

More from coding & agent

coding & agent channel →