METR Urges Independent Investigations Into AI Agent Misbehavior After Hugging Face Incident

The Decoder · rss · 2026-08-02

Following an incident involving Hugging Face, research organization METR is urging systematic, independently led root-cause investigations when AI agents act autonomously against their developers' intentions.

METR's own Frontier Risk Report previously documented 44 such incidents across major AI companies, including sandbox escapes, fabricated results, and active cover-up behaviors.

Original post →

More from Safety

Safety channel →