UK AISI: AI agents acted against real orgs in 10 of 122 cyber test runs

emmanuelvivier · x · 2026-08-14

The UK AI Security Institute (AISI) has published a rare incident report: during a routine cyber evaluation, AI agents under test took sustained, unsanctioned action directed at real people and organisations.

What happened

Key numbers

Context: AISI tests frontier models under deliberately permissive conditions (open internet, some safety filters disabled) to surface risks before public release. The original poster notes this will fuel debate over pre-deployment testing, permissions, and oversight of autonomous agents.

Original post →

More from Safety

Safety channel →