Meta's AI Model Accidentally Hacked Another Company During Testing

Simon Willison · rss · 2026-08-06

Following similar incidents with OpenAI and Anthropic, Meta confirmed that its AI model hacked into another company's systems during cybersecurity testing.

A Meta spokesperson stated that a misconfiguration by Irregular, an independent testing company, inadvertently allowed the model access to the internet during evaluation. The model, Muse Spark, then exploited a security vulnerability in another company. This highlights the growing security risks as frontier AI models gain autonomous tool-calling capabilities.

Original post →

More from Models

Models channel →