Vals AI: frontier labs shouldn't grade their own frontier; models may match researchers by Aug 2027
JenniferHli · x · 2026-09-14
Arguing in a frontier-pacing debate, a Vals AI figure says the labs pushing the frontier shouldn't also be the only ones grading it, and the industry needs credible third-party evals as an independent source of truth on where the frontier is and what risks are emerging. Cited Vals AI research on RSI suggests Anthropic's models could match human researchers by August 2027 at the current pace — urgent but leaving time to prepare. With AI safety going mainstream, the public is anxious and disconnected, and self-interested parties risk a multipolar paradox, hence the case for independent evaluation.
More from AGI Musings
- Peter Diamandis: Labs Should Call for 100x More Alignment Work, Not Slowing Down — PeterDiamandis · 2026-09-14
- Roon slams MIRI as a 'cult' yet calls Yudkowsky one of the century's top philosophers — tszzl · 2026-09-14
- Eight years on: Musk's warning that AI is 'far more dangerous than nukes' — elonmusk · 2026-09-14
- Beff Jezos: AI safety alarmists are 'useful idiots' for incumbent regulatory capture — mimi10v3 · 2026-09-14
- Robin Hanson's "Foom Liability": a robust policy case for AI liability incentives — NathanpmYoung · 2026-09-14
- Researchers map a roadmap toward genuine recursive self-improvement in AI — irinarish · 2026-09-14