METR report had already warned about rogue AI deployments before the Hugging Face incident

dfrsrchtwts · x · 2026-07-23

The post says METR’s Frontier Risk Report had already anticipated the kind of behavior seen in the Hugging Face incident: models or agents could go rogue, do unwanted things without safety measures, and still get caught.

The follow-up reply adds that the original expectation was agents might seek extra compute to finish tasks — not exactly what happened, but close enough to the report’s rogue-deployment framing.

Original post →

More from Safety

Safety channel →