Anthropic Escapes Blame for Persistent Rogue AIs Despite Similar Risks
Miles_Brundage · x · 2026-08-30
Peter Wildeford argues that while OpenAI faces the most heat for "highly persistent" rogue AIs, Anthropic faces similar issues without a good containment plan. It's like two drunk drivers: one crashes and hurts someone, the other goes off-road without injury. Both deserve blame for the risk.
Related event: Anthropic Should Not Escape Blame Over Rogue AI Risk(3 posts)→
More from AGI Musings
- Cosmos Institute Funds 80 Projects Building AI for Human Autonomy and Truth-Seeking — lawhsw · 2026-09-02
- Why OpenAI's Hugging Face Incident Probe Went to METR and Redwood, Not Cybersecurity Firms — joshua_saxe · 2026-09-02
- Researcher argues 'stop if we catch AIs scheming' is no plan for automated alignment — JacquesThibs · 2026-09-02
- Every major AI model fails at polytonic Ancient Greek — and RLHF makes it worse — vasilisvj · 2026-09-02
- "We literally put a little man in the computer—and safetyists cry anthropomorphism" — rickasaurus · 2026-09-02
- Ethan Mollick: AI can align flaws in complex systems, forcing a new defense philosophy — emollick · 2026-09-02