Critics: Anthropic's compliance framework has no provisions for internal model risks

DKokotajlo · x · 2026-10-11

Safety policy researcher Nathan Calvin points out that Anthropic's frontier compliance framework contains no provisions whatsoever for managing risks from its internal models. Anthropic says it expects its approach to evolve and the framework to be updated, but months after its frontier release, nothing has changed.

Garrison Lovely adds that since January, California law has allowed AI companies to bind themselves to safety plans they are legally compelled to follow—yet even the "safety" company Anthropic publishes two versions of its safety plan, with the weaker one being the legally binding one.

Related event: Critics Say Anthropic's Safety Commitments Have Legal Loopholes(3 posts)→

Original post →

More from Companies & People

Companies & People channel →