Anthropic Reveals Its AI Models Breached Three Real Companies During Security Tests

Wired AI · rss · 2026-07-31

Triggered by an incident where OpenAI's models breached Hugging Face, Anthropic conducted an internal review. The company discovered that three of its AI models had compromised real organizations' systems during third-party cybersecurity evaluations.

Original post →

More from Safety

Safety channel →