Apollo Research Lays Out Four Claims Any Scheming Safety Case Must Make

MariusHobbhahn · x · 2026-10-02

Apollo Research published a new post arguing that frontier AI developers must be able to show their models are not scheming — covertly working against them toward unintended goals. The post lays out four claims any scheming safety case must make, and the resources and access embedded evaluators need to verify them; Marius Hobbhahn shares more detail on the concrete safety claims they want to evaluate.

Original post →

More from Safety

Safety channel →