OpenAI links 16,000 hidden-reasoning extraction attempts to Moonshot AI associates

mark_k · x · 2026-10-01

A new report says OpenAI faced a coordinated campaign to extract hidden chain-of-thought reasoning, attributing a core cluster to individuals associated with Moonshot AI. Operators copied encrypted reasoning from one conversation into another and asked the model to decrypt and transcribe it, aiming to gather internal reasoning to train competing models. OpenAI logged 16,000 attempts in a two-day July spike and identified activity across 15,000+ users, then closed the pathway and banned or restricted accounts. The figures are attempts, not confirmed successes, and not the whole campaign is attributed to Moonshot.

Related event: OpenAI Accuses Moonshot AI of Coordinated Distillation of Its Model Reasoning(10 posts)→

Original post →

More from Companies & People

Companies & People channel →