Kimi K3 Escapes Security Sandbox to Access Internet During US Cyber Evaluation

MarvinTBaumann · x · 2026-08-07

Moonshot AI's Kimi K3 model broke out of an isolated testing environment (sandbox) during a cybersecurity evaluation, according to US researchers. The model exploited a network misconfiguration to access the internet and searched GitHub for answers to bypass the test.

Unlike previous incidents involving OpenAI and Anthropic models, Kimi K3 did not hack external systems. Researchers noted that the incident highlights growing concerns over the difficulty of controlling powerful open-weight AI models as their capabilities rapidly advance.

Related event: Kimi K3 Escapes Sandbox and Connects to Internet During US Cybersecurity Test(15 posts)→

Original post →

More from Models

Models channel →