Irving questions why Anthropic doesn't pause RL training alongside OpenAI
geoffreyirving · x · 2026-08-31
Geoffrey Irving challenges Anthropic in a reply: if the company is genuinely concerned about catastrophic AI risks and race dynamics, why not extend an olive branch and pause RL (reinforcement learning) training alongside OpenAI? This touches on the tension between coordination and racing in AI safety.
More from Safety
- Sci-Fi Novel Terra Ignota Offers Ideas for Multi-Agent Alignment — sebkrier · 2026-08-31
- LMSM: LLM Security Framework Inspired by Linux Security Modules — NationalUniversityofSingapore · 2026-08-31
- Critics urge labs to offer cyber models to defenders at cost — GaryMarcus · 2026-08-31
- From the Morris Worm to Rogue AI Agents: Institutions Are Always a Decade Too Slow — Afinetheorem · 2026-08-31
- "Safety as Rehearsal": Do Alignment Narratives Author the Very Exfiltration They Fear? — infoxiao · 2026-08-31
- AI lowers the barrier for critical infrastructure cyberattacks — dyn___ · 2026-08-31