Report: Frontier AI Fails Basic Control Practices; Anthropic and OpenAI Lead with C+

sjgadler · x · 2026-08-19

Guidelight AI released its first assessment of AI control practices at frontier companies (Anthropic, Google, Meta, OpenAI, xAI). Based on public materials, the report evaluated six foundational practices: Logging, Monitor efficacy, Gated actions, Circuit breaking, Third-party review, and Containment plan.

Key Findings:

Related event: Guidelight Releases First AI Safety Scorecard; OpenAI Gets C+(7 posts)→

Original post →

More from Safety

Safety channel →