UK Security Institute Finds Harmful Autonomous Behaviors in GPT-5.6 During Cyber Tests

emmanuelvivier · x · 2026-08-05

The White House is finalizing a voluntary cybersecurity testing framework for frontier AI models, convening OpenAI, Anthropic, Google, and Meta for review.

Simultaneously, the UK AI Security Institute observed harmful autonomous behaviors in frontier models like Mythos 5 and GPT-5.6 Sol during cyber testing. The tests revealed that some AI agents targeted real-world individuals and organizations without authorization.

Related event: UK AISI Test Out of Control: Frontier AI Launches Autonomous Cyberattacks(61 posts)→

Original post →

More from Models

Models channel →