Why unattended agents redo finished work: a thought experiment on task completion state

ClickOk5811 · reddit · 2026-10-09

A Reddit post frames a thought experiment for a common long-run agent failure mode: imagine a coworker who can never cross finished tasks off a list of 40 jobs — by job 25, they'd re-examine and likely redo completed work.

That, the author argues, is what an unattended agent does on long multi-step runs, like a refactor that renames an interface and updates every call site, then re-touches files it already got right hours later. Nothing marks steps as "closed," and every finished tool call sits in context weighted like the current step.

The problem shows up less with a human in the loop because each reply implicitly tells the model prior steps are settled; remove the human and nothing takes over that role. A bigger context window won't fix it — it's just more room for finished work to linger. The author points to explicit mechanisms: a task list the agent must update before moving on, or collapsing finished steps into one-line notes. Untested; the author asks overnight-job runners what has worked for them.

Original post →

More from coding & agent

coding & agent channel →