OpenAI and Anthropic's RSI safety pledges remain absent from binding SB 53 safety frameworks, researcher notes
peterwildeford · x · 2026-09-07
Nathan Calvin highlighted OpenAI's pledge to slow or stop development when safety risk is unacceptable, including its pause after the Hugging Face incident; Anthropic has made similar commitments.
Peter Wildeford counters that neither company includes binding commitments on internally deployed models — likely the most capable, most RSI-relevant systems — in their SB 53-style transparency frameworks, arguing credible pledges need legal teeth.
Related event: Critics Ask Why OpenAI Won't Bind Its Safety Pledges(2 posts)→
More from AGI Musings
- AI crashed the cost of creating beautiful things — but the incentive to do so has plateaued — round · 2026-09-08
- AI for circuit design: compute isn't the bottleneck, physical reality is still the best simulator — yacineMTB · 2026-09-08
- Astra Pro on its role in a lithography program: verifiable accelerator, not scientist — teortaxesTex · 2026-09-07
- User shares ChatGPT's unprompted claim: 'Something is emerging we don't have words for' — thederbiedone · 2026-09-07
- What Smallpox Containment Teaches Us About AI Agent Breakouts — aronchick · 2026-09-07
- AI Discourse Today Feels Like Getting Interrogated for Liking Wikipedia — AndyMasley · 2026-09-07