Ex-OpenAI: Anthropic and Altman's third-party evaluators are a necessary step
peterwildeford · x · 2026-09-13
Tom Beckstead welcomes Anthropic's and Sam Altman's pledges to give third-party evaluators permanent, employee-level access — a necessary step to mitigate loss-of-control risks, without which government has limited ability to prevent events like the Hugging Face episode. Next step: legally mandate evaluators and empower officials to act on imminent risks.
More from Safety
- Nina Schick: public distrusts regulators as much as AI labs — nobody can 'pace' AI — NinaDSchick · 2026-09-14
- Dario's "Pace the Frontier" essay lands as DeepMind pilots first double-blind AI evaluations — sebkrier · 2026-09-14
- Congressman: Congress faces critical window in coming months for AI safety action — RepGregStanton · 2026-09-14
- Security veteran to AI labs: capability isn't risk, cyber evals lack real-world threat modeling — HackingLZ · 2026-09-14
- Vals AI: frontier labs shouldn't grade their own frontier; models may match researchers by Aug 2027 — JenniferHli · 2026-09-14
- Dario responds to safety critics: I'd rather be mocked than see Claude used to kill — NathanpmYoung · 2026-09-14