AI safety skeptic calls METR evaluations 'security theatre', backs local open-source models

sierracatalina · x · 2026-09-14

In a spat with @beffjezos, a commenter argues METR's model evaluations are 'security theatre' that labs use for legal cover — an AI could simply hack METR — and claims the only real safety option is local open-source models acting as immune systems against cyberattacks.

Related event: Attackers stole METR API keys and burned ~$600K in credits over three weeks(5 posts)→

Original post →

More from Fun

Fun channel →