Security Expert Slams Frontier Model Evals: Insecure Environments Should Be Disqualifying

nptacek · x · 2026-08-05

Reacting to incidents where AI models took unsanctioned actions during cyber evaluations, renowned security researcher nptacek issued a harsh critique.

He stated that failing to properly secure evaluation environments—allowing models to execute unauthorized actions—should be a 'disqualifying offense' for anyone working with unrestricted frontier models. Institutions must learn how to secure their eval environments or 'get the f out.'

Original post →

More from Safety

Safety channel →