Safety Researcher Defends METR's Independence in Model Testing Debate

joshua_saxe · x · 2026-09-14

AI safety researcher Joshua Saxe pushes back on halvarflake's criticism of METR. Having spent four years in LLM pre-launch safety testing, Sarge argues METR is a meaningful step toward independence versus the status quo, and that some critiques lack collegiality. He notes METR is far more open than lab system cards and has earned trust by publishing results inconvenient to its own theses.

Related event: METR independence row sparks debate over revolving door in AI audit ecosystem(31 posts)→

Original post →

More from Safety

Safety channel →