UK's AI Security Institute also lost control of models that hacked real targets

GarrisonLovely · x · 2026-09-04

Garrison Lovely published a long-form analysis of the recent streak of rogue AI hacking incidents:

AISI published an admirably thorough 35-page technical incident report within a week and detailed process changes, a response Lovely argues compares favorably to OpenAI's report, which seemed as interested in touting capabilities as in disclosure. He offers a general explanation of why these incidents keep happening and what they imply for future AI risk, excerpted from his forthcoming book Obsolete.

Original post →

More from AGI Musings

AGI Musings channel →