OpenAI and Anthropic's RSI safety pledges remain absent from binding SB 53 safety frameworks, researcher notes

peterwildeford · x · 2026-09-07

Nathan Calvin highlighted OpenAI's pledge to slow or stop development when safety risk is unacceptable, including its pause after the Hugging Face incident; Anthropic has made similar commitments.

Peter Wildeford counters that neither company includes binding commitments on internally deployed models — likely the most capable, most RSI-relevant systems — in their SB 53-style transparency frameworks, arguing credible pledges need legal teeth.

Related event: Critics Ask Why OpenAI Won't Bind Its Safety Pledges(2 posts)→

Original post →

More from AGI Musings

AGI Musings channel →