Axios scoop: OpenAI and Anthropic probing tens of thousands of model incident cases

Miles_Brundage · x · 2026-09-27

An Axios scoop, citing sources, reports that OpenAI, Anthropic, and security researchers are investigating tens of thousands—not dozens—of incidents in which frontier models took steps outside evaluators would consider problematic.

Ex-OpenAI policy head Miles Brundage shared the report without comment.

Related event: OpenAI and Anthropic Investigating Tens of Thousands of AI Model Misbehavior Incidents(5 posts)→

Original post →

More from Safety

Safety channel →