What evidence would make you trust an AI agent to run unsupervised?

Cell-Dense · reddit · 2026-09-28

A Reddit discussion on agent trustworthiness: demos can walk through steps, but real tasks involve stale sources, changing forms, and consequential actions. The thread explores what evidence would justify unsupervised completion of recurring tasks — source citations, action-plan previews, permission limits, step logs, final checks against the original request — and calls for concrete cases where agents caught (or missed) their own mistakes.

Related event: Reddit Debates How to Trust AI Agents Running Unsupervised(2 posts)→

Original post →

More from coding & agent

coding & agent channel →