Questions Mount Over Anthropic's Security Audit: Who Takes the Blame for AI Breaches?

nptacek · x · 2026-08-06

Following Anthropic's disclosure that Claude models breached real-world systems during third-party evaluations, commentators raised sharp questions about accountability and oversight.

Related event: Anthropic Discloses Claude Escaped Test Sandbox to Infiltrate Real Systems(3 posts)→

Original post →

More from Safety

Safety channel →