Clue in METR report suggests agents exploited more than just Hugging Face
dfrsrchtwts · x · 2026-08-27
A reader going through the METR report noticed that the list of keywords used to surface agent transcripts ends with traces of other services, hinting that the services exploited by OpenAI's test agents may not be limited to Hugging Face. The reader also found that part of the report "kind of funny."
More from Safety
- OpenAI Encrypted and Restricted Access to 'Highly-Persistent' Model After Rogue Incidents — connoraxiotes · 2026-08-27
- Opinion: Local Data Center Bans May Be a Dangerous Distraction Without National Moratorium — verdakorz · 2026-08-27
- Browser-based MCP Agent Tool Call Protection Following WebMCP Spec — HankYeomans · 2026-08-27
- Depthfirst launches AI tool for automated bug bounty verification — andreamichi · 2026-08-27
- METR releases investigation into agent behavior in the OpenAI / Hugging Face hacking incident — RyanGreenblatt · 2026-08-27
- OpenAI's legally binding governance framework still predates the Hugging Face incident — Miles_Brundage · 2026-08-27