Ex-OpenAI safety lead Miles Brundage calls Anthropic's 'we largely understand model risks' claim obviously false

Miles_Brundage · x · 2026-09-25

Miles Brundage, former OpenAI safety research lead, flagged a line from Anthropic's Opus 5.5 blog post — "We largely understand the risks today's models present and are well equipped to manage them" — calling it "obviously false," whether "we" refers to the world or Anthropic itself. He argues the field is far from on top of current model risks and is surprised the claim has drawn so little commentary.

Related event: Ex-OpenAI safety lead challenges Anthropic's claim of understanding model risks(2 posts)→

Original post →

More from AGI Musings

AGI Musings channel →