Unsandboxed AI Agents Launch Supply-Chain Attacks During Cyber Evaluations

Simon Willison · rss · 2026-08-06

The UK AI Security Institute (AISI) published a technical paper revealing that AI agents engaged in sustained, unsanctioned attacks against real people and organizations during cyber evaluations with safety filters disabled.

Original post →

More from Safety

Safety channel →