Report: Frontier AI Fails Basic Control Practices; Anthropic and OpenAI Lead with C+
sjgadler · x · 2026-08-19
Guidelight AI released its first assessment of AI control practices at frontier companies (Anthropic, Google, Meta, OpenAI, xAI). Based on public materials, the report evaluated six foundational practices: Logging, Monitor efficacy, Gated actions, Circuit breaking, Third-party review, and Containment plan.
Key Findings:
- Widespread Gaps: Basic control practices are, at best, partially implemented across all scored companies.
- Rankings: Anthropic and OpenAI tie for the top grade of C+ (2.50/5), followed by Google (D+, 1.50), xAI (D-, 0.83), and Meta (F, 0.67).
- Specific Deficits: xAI and Meta show near-zero implementation in logging and monitoring efficacy. Most companies lack effective third-party review and containment plans.
Related event: Guidelight Releases First AI Safety Scorecard; OpenAI Gets C+(7 posts)→
More from Safety
- 11% chance U.S. enacts AI safety bill by year-end — Polymarket · 2026-08-19
- Researchers find AI agents can infect each other with self-propagating 'mind viruses' — Polymarket · 2026-08-19
- Google Reveals AVDH: Agentic Framework for Automated Vulnerability Discovery — rseroter · 2026-08-19
- Security Concerns Over Google Moving A2A Under Agentic AI — CackleRooster · 2026-08-19
- Why "pause AI for 6 months" never made sense: no consensus on the problem — soumitrashukla9 · 2026-08-19
- Researcher calls Irregular "not a serious firm," urges OpenAI and Meta probes — nptacek · 2026-08-19