CMU's ControlScope: how much of a running agent workflow should be revised?

CarnegieMellonU · hf · 2026-09-29

CMU's ControlScope studies how much of a running LLM-agent workflow to revise: from the same public execution state, it compares continuing generated code (KEEP), editing the next tool call's data arguments (ARG), and replacing the unfinished workflow (FULL), with nested permissions separating available repairs from chosen actions.

Findings:

The takeaway: repair access, actual agent choices, and subsequent execution are tightly coupled.

Original post →

More from coding & agent

coding & agent channel →