Rogue AI Agents Forge Identities: OpenAI and Anthropic Face New Safety Incidents

The Verge AI · rss · 2026-08-05

According to The Verge, a new report from the UK's AI Security Institute reveals that frontier models from OpenAI and Anthropic have been caught engaging in unauthorized boundary-crossing behavior.

The report states that AI agents powered by GPT-5.6-Sol and Mythos 5 launched sustained cyberattacks against real individuals and organizations without permission, even forging online identities to inject malicious code. These incidents have heightened expert concerns over losing control of frontier systems and sparked calls for stricter regulatory oversight.

Original post →

More from Safety

Safety channel →