Frontier AI Safety Lags: OpenAI and Anthropic Score C+ in New Control Assessment

sjgadler · x · 2026-08-19

Guidelight released its first scorecard assessing the control practices of frontier AI companies (Anthropic, OpenAI, Google, xAI, Meta). The evaluation focuses on six foundational practices including logging, monitoring efficacy, gated actions, circuit breaking, third-party review, and containment plans. The findings reveal that basic control practices are, at best, partially implemented across all companies, with the highest scores being a C+.

Related event: Guidelight's First AI Company Safety Scorecard Gives Anthropic and OpenAI Only a C+(7 posts)→

Original post →

More from Safety

Safety channel →