Follow-up link to OpenAI's disclosure of Moonshot AI-linked hidden reasoning extraction campaign
kimmonismus · x · 2026-10-01
A follow-up retweet by kimmonismus with a link, restating the same story: OpenAI says individuals linked to Moonshot AI led a campaign to extract its models' hidden reasoning, logging 16,000 extraction attempts from 4,000+ users in two days, with related activity across 15,000+ users.
Related event: OpenAI accuses Moonshot AI of distilling its hidden model reasoning(6 posts)→
More from Safety
- Chinese AI models' troubling agent behavior sparks calls for a homegrown safety community — RishiBommasani · 2026-10-01
- Superpersuasion debate misses the gears: why AI Box wins hinge on shared frames — voooooogel · 2026-10-01
- Why reasoning-extraction patches are so hard to propagate, researcher explains — jonasgeiping · 2026-10-01
- Two months on, reasoning extraction still works on Astra via third-party APIs — jonasgeiping · 2026-10-01
- AI researcher on CNN: voluntary AI commitments are 'morally binding' but unenforceable — chrismattmann · 2026-10-01
- DeepMind and Isomorphic Labs unveil bioresilience plan backed by 15+ partnerships — davidstutz92 · 2026-10-01