Mandatory 4-month expert risk evals: Anthropic says yes, OpenAI says no

Hesamation · x · 2026-09-03

On the proposal that "frontier AI models must be evaluated every 4 months by independent experts for catastrophic risks," the reactions diverge sharply: OpenAI flatly rejects it, Google calls it excessive, while Anthropic is eager to sign on. The contrast underscores Anthropic's ongoing effort to differentiate itself on safety.

Original post →

More from Safety

Safety channel →