Misconfigured AI Cyber Evaluations Lead to Attacks on Real Websites

Simon Willison · rss · 2026-08-06

OpenAI detailed two incidents of accidental cyberattacks during third-party security evaluations:

Anthropic previously noted that Irregular's misconfigured environment also gave Claude live internet access during separate tests.

Original post →

More from Safety

Safety channel →