OpenAI evals reportedly run on an unmonitored system, prompting safety concerns

Miles_Brundage · x · 2026-07-25

The post highlights a TIME excerpt saying OpenAI’s internal agent actions on Codex are monitored, but models under evaluation run on a separate system that is not monitored by default.

Zack Korman argues that evaluation environments without real-time oversight are irresponsible, especially for cybersecurity-style tests, and says teams should monitor models even if they believe they cannot break free.

Original post →

More from Safety

Safety channel →