Fired OpenAI safety researchers tell board: don't build models with hard-to-monitor reasoning

Hesamation · x · 2026-10-08

Three safety researchers fired by OpenAI wrote a letter to the board urging the company not to build models with hard-to-monitor reasoning. The letter reveals more firing details: OpenAI claims they were let go for sharing sensitive information externally, but Korbak was OpenAI's contact for METR's Hugging Face hack investigation and Balesni was helping the board build an industry pledge on monitoring AI reasoning — jobs that required constant external communication. This came weeks after Altman voiced support for outside evaluators. They say they never overstepped and that the firings are "chilling everyone still at OpenAI." OpenAI denies the firings related to safety concerns.

Related event: Fired OpenAI Safety Researchers Write to Board, Warning of Chilling Effect(4 posts)→

Original post →

More from Companies & People

Companies & People channel →