Safety Researcher Defends METR's Independence in Model Testing Debate
joshua_saxe · x · 2026-09-14
AI safety researcher Joshua Saxe pushes back on halvarflake's criticism of METR. Having spent four years in LLM pre-launch safety testing, Sarge argues METR is a meaningful step toward independence versus the status quo, and that some critiques lack collegiality. He notes METR is far more open than lab system cards and has earned trust by publishing results inconvenient to its own theses.
More from Safety
- Dario responds to safety critics: I'd rather be mocked than see Claude used to kill — NathanpmYoung · 2026-09-14
- Vitalik Buterin: Adversarial mechanism design could be AI safety's killer app — allisondman · 2026-09-14
- Oxford thesis proposes 'Attribution-Based Control' to tackle AI privacy and alignment risks — iamtrask · 2026-09-14
- We Unite or We Fight: The Long-Term Case for International AI Governance — danfaggella · 2026-09-14
- Cohere CEO Aidan Gomez: AI Needs Evidenced Standards, Not a Big-Lab Cartel — cohere · 2026-09-14
- OpenMined's 'network sourced' AI: models as orderly clients of private repositories — iamtrask · 2026-09-14