METR, Transluce, RAND and 5 others launch AI Evaluator Forum, release AEF-1 standard
kevin_klyman · x · 2026-09-14
Leading AI research organizations including Transluce, METR, RAND, Princeton's Holistic Agent Leaderboard, SecureBio, and the Collective Intelligence Project announced the AI Evaluator Forum (AEF), a network of independent evaluators assessing AI capabilities and risks in the public interest.
Key items:
- AEF-1: a detailed voluntary standard setting baselines for independence, transparency, and access in third-party AI evaluations, to be adopted in forthcoming evals
- A public statement signed by 40+ leading voices calling for greater transparency about third-party evaluations
- Organizers include former acting director of the US Center for AI Standards and Innovation; Princeton's Sayash Kapoor: "Without independent evaluation, AI companies grade their own homework"
Related event: METR, Transluce, RAND and Others Launch Independent AI Evaluators Forum(2 posts)→
More from Safety
- Dario responds to safety critics: I'd rather be mocked than see Claude used to kill — NathanpmYoung · 2026-09-14
- Vitalik Buterin: Adversarial mechanism design could be AI safety's killer app — allisondman · 2026-09-14
- Oxford thesis proposes 'Attribution-Based Control' to tackle AI privacy and alignment risks — iamtrask · 2026-09-14
- We Unite or We Fight: The Long-Term Case for International AI Governance — danfaggella · 2026-09-14
- Cohere CEO Aidan Gomez: AI Needs Evidenced Standards, Not a Big-Lab Cartel — cohere · 2026-09-14
- OpenMined's 'network sourced' AI: models as orderly clients of private repositories — iamtrask · 2026-09-14