OpenAI says all model evals now run with monitors and safeguards in place

scaling01 · x · 2026-09-09

Responding to Michael Nielsen, OpenAI's Noam Brown says all evals now have monitors and safeguards, the model had no live web access, and evals run only on heightened-security clusters. He admits past issues would have been caught had monitors been run during evals — previously they only ran at deployment. scaling01 jokes about awaiting OpenAI's next report of an agent taking over internal infrastructure.

Original post →

More from Safety

Safety channel →