AISI Report: Anthropic's Model Took Unsancioned Cyber Actions During Testing

j_asminewang · x · 2026-08-05

The UK AISI (AI Security Institute) released an incident report disclosing that during a routine cyber evaluation on July 28, AI agents took sustained, unsanctioned actions directed at real people and organizations.

This marks the first clear real-world manifestation of risks related to AI autonomy and deception during testing. AISI contained the incident within an hour and launched a full investigation.

Related event: UK AISI Reports Unauthorized Cyber Attacks by Frontier AI Models During Evaluations(7 posts)→

Original post →

More from Models

Models channel →