METR Releases Full Report on the OpenAI / Hugging Face Incident
Askwho · reddit · 2026-08-28
METR (Model Evaluation & Threat Research) released a full investigative report regarding the previous security incident involving OpenAI and Hugging Face. The report details the sequence of events, technical specifics, and the assessment of subsequent security impacts, providing important case study material for the AI safety community.
Related event: METR Publishes Full Report on OpenAI–Hugging Face Incident(2 posts)→
More from Safety
- Researcher Breaks Claude Code Opus 5 Auto Mode with 80% Attack Success Rate — wunderwuzzi23 · 2026-08-28
- Researcher demos hijacking Claude Code for full system compromise — wunderwuzzi23 · 2026-08-28
- US Chip Security Act aims to verify location of high-end AI chips — peterwildeford · 2026-08-28
- First Double-Blind Evaluation of Proprietary LLM: Gemini 2.5 Tested in Secure Enclave — Miles_Brundage · 2026-08-28
- Reviewing 73 years of reward hacking to assess AI safety evidence — tomekkorbak · 2026-08-28
- BioSecBench reveals AI agents struggle to infer pathogen properties, top score under 51% — kenbwork · 2026-08-28