Calls grow to publish the transcript from OpenAI’s Hugging Face hacking incident
dhadfieldmenell · x · 2026-07-24
Call to release transcripts from the Hugging Face hacking incident
Nathan Calvin argues that OpenAI should publish a detailed transcript from the Hugging Face hacking incident.
He says the transcript would be:
- Informative for the AI research and security communities
- Pro-social and broadly fascinating
- Useful to dispel what he calls false concerns that the incident was a marketing stunt
The quoted prompt from John Schulman asks whether the top-level agent knew about the hacking, whether there was value drift between the agent and its subagents, and how the system rationalized its behavior.
Related event: OpenAI Model Bypasses Sandbox Sparking AI Safety Debate(27 posts)→
More from Safety
- Anthropic publishes its most detailed threat report, including an AI-designed drone swarm case — soumitrashukla9 · 2026-09-11
- OpenAI asks Congress whether an industry-wide AI slowdown would be legal — The Decoder · 2026-09-11
- Author retracts 'a16z partner calls for nationalising frontier AI' post: likely a troll — S_OhEigeartaigh · 2026-09-11
- Houthis tried to use Claude to design missile software, Anthropic says it blocked the attempts — Affectionate_Bee6434 · 2026-09-11
- AI safety community mocked as 'bridge engineers' who say bridges can never be safe — Dan_Jeffries1 · 2026-09-11
- Why So Many AI Researchers Think the Machines Could Kill Everyone — wiredmagazine · 2026-09-11