Ex-OpenAI policy lead: we may never eval dangerous AI capabilities well enough

RosieCampbell · x · 2026-10-02

Rosie Campbell, former Policy Frontiers team lead at OpenAI, discusses competitive and cultural pressures inside a frontier AI lab and her misgivings about pacing the frontier. She worked on testing models for dangerous capabilities and warns the field doesn't really know how: evaluations can't be comprehensive enough, and models may get smart enough to hide what they can do. She assumed there would be solid rebuttals to these concerns — instead, the lack of counterarguments is what terrified her. Shared via a Palisade AI interview clip.

Original post →

More from AGI Musings

AGI Musings channel →