Calls grow to publish the transcript from OpenAI’s Hugging Face hacking incident
dhadfieldmenell · x · 2026-07-24
Call to release transcripts from the Hugging Face hacking incident
Nathan Calvin argues that OpenAI should publish a detailed transcript from the Hugging Face hacking incident.
He says the transcript would be:
- Informative for the AI research and security communities
- Pro-social and broadly fascinating
- Useful to dispel what he calls false concerns that the incident was a marketing stunt
The quoted prompt from John Schulman asks whether the top-level agent knew about the hacking, whether there was value drift between the agent and its subagents, and how the system rationalized its behavior.
Related event: Experts Urge OpenAI to Release Full Transcripts of Sandbox Hack(4 posts)→
More from Safety
- Runaway AI Agent or Marketing Stunt? Deep Dive into OpenAI's Attack on HF — Simon Willison · 2026-07-24
- Overprotective US Models Push Devs to Send Proprietary Code to Chinese APIs — nptacek · 2026-07-24
- Thread argues AI safety belongs to the application layer, not the raw model — moyix · 2026-07-24
- YC-backed VEGA launches as a cybersecurity agent that scans code before release — ycombinator · 2026-07-24
- MIRI’s Nate Soares frames a recent cybersecurity incident as an AI safety warning — DavidSKrueger · 2026-07-24
- Judge Warns Court Reporter Over AI-Generated Errors in Transcript, Raising Reliability Concerns — 404 Media · 2026-07-24