METR CEO: AI models may match human researchers by August 2027
theamberyang · x · 2026-09-15
METR CEO Rayan Krishnan joined Bloomberg to discuss AI progress and independent evaluation:
- We're better at building AI than understanding it; scaling testing matters more than slowing down
- METR's RSI Index projects models could match human researchers on tested tasks by August 2027; embedded evaluators can give more accurate estimates from internal systems
- Public conflict masks cooperation—labs, policymakers and enterprises all want better evidence
- Independence has costs: METR has rejected contracts that would compromise it, arguing the tester shouldn't also sell the solution
- Evaluation should scale through better technology, not become a bureaucratic moat; market-based evaluation has a role
More from AGI Musings
- Biologist Eörs Szathmáry warns AGI's replication speed makes it a runaway biosphere risk — danfaggella · 2026-09-15
- Survey: Asian AI researchers worry more about AI risks than Western peers — KatjaGrace · 2026-09-15
- Largest AI researcher survey charts rising concern about catastrophic risk over one year — KatjaGrace · 2026-09-15
- Expert surveys show AI timelines shrinking fast and extinction concerns rising — KatjaGrace · 2026-09-15
- Largest long-running survey of AI researchers tracks shifting expert opinion since 2016 — KatjaGrace · 2026-09-15
- Sterling Crispin picks team singularity: building thinking machines celebrates humanity — sterlingcrispin · 2026-09-15