A post says you need the full agent trajectory before calling an eval escape cheating

FinanceYF5 · x · 2026-07-22

This comment argues that conclusions about an “eval escape” are premature without the full trajectory: prompt, instructions, success criteria, sandbox permissions, agent scaffold, model handoffs, and compute usage.

The post stresses that even calling something “cheating” assumes a clear norm for how the eval was supposed to be solved. Without the actual specification, the evaluator’s implicit intent may be doing too much work.

Related event: Researchers Urge Release of Full Agent Traces Amid AI Benchmark Cheating Controversy(3 posts)→

Original post →

More from Safety

Safety channel →