Vals AI: frontier labs shouldn't grade their own frontier; models may match researchers by Aug 2027

JenniferHli · x · 2026-09-14

Arguing in a frontier-pacing debate, a Vals AI figure says the labs pushing the frontier shouldn't also be the only ones grading it, and the industry needs credible third-party evals as an independent source of truth on where the frontier is and what risks are emerging. Cited Vals AI research on RSI suggests Anthropic's models could match human researchers by August 2027 at the current pace — urgent but leaving time to prepare. With AI safety going mainstream, the public is anxious and disconnected, and self-interested parties risk a multipolar paradox, hence the case for independent evaluation.

Original post →

More from AGI Musings

AGI Musings channel →