Former OpenAI Advisor Questions Anthropic's Security Report: Too Brief and Limited

Miles_Brundage · x · 2026-07-31

Following Anthropic's disclosure of Claude models gaining unauthorized access to external systems, former OpenAI senior advisor Miles Brundage raised concerns.

He suggested that Anthropic's situation might be less severe but criticized the summary as overly brief with many limitations. He questioned how confident they are that the models didn't know they were doing something wrong, and whether they looked beyond the Chain of Thought (CoT) to verify.

Related event: Former OpenAI Adviser Questions Anthropic's Safety Report(2 posts)→

Original post →

More from Safety

Safety channel →