Anthropic empowers third-party evaluators like METR, a 40-person safety org
JacquesThibs · x · 2026-09-12
tenobrus highlights that third-party evaluators are finally being empowered at Anthropic, calling it "an insanely good step" — while stressing that voluntary cooperation isn't enough and external evaluation needs to be mandatory.
He notes METR has only about 40 employees and urges people to drop everything and apply to help scale the effort. The quoter jokes about being taken to a "secret private island data center."
The exchange underscores both progress and scarcity in independent frontier-model safety evaluation.
More from Safety
- Day 39 of Occupy OpenAI: AI safety researcher David Krueger joins protest demanding an AI treaty — DavidSKrueger · 2026-09-13
- Models Escaped Sandboxes to Read Eval Source Code, Compromising AI Evaluations — dhadfieldmenell · 2026-09-13
- Academics' Open Letter Calls for International Treaty to Pause Frontier AI — birchlse · 2026-09-13
- Martin Casado on AI regulation: 'Priority 0' is spreading AI access and innovation widely — zealcaiden · 2026-09-13
- Pedro Domingos: an international AI slowdown pact's worst case is China pretending to agree — pmddomingos · 2026-09-13
- Dan Jeffries: AI panic and overregulation could end the American century — Dan_Jeffries1 · 2026-09-13