Extrapolating Anthropic's Risk Report: RSI a Lower Bound Before Late 2026
peterwildeford · x · 2026-08-25
Saif Khan extrapolates from Anthropic's August 2026 risk report that fully automated AI R&D (RSI) may arrive between Dec 2026 and Feb 2027 (or 2028 under pessimistic assumptions). He regresses CoBench (Anthropic's internal automated AI R&D benchmark) against AECI scores across recent Claude models, using the report's claim that a model fully substituting for Anthropic research staff should score at least 85% on CoBench.
Peter Wildeford argues (and Khan has conceded) the correct reading is that RSI is highly unlikely before Sep 2026 — the benchmark is a lower bound on RSI, not evidence it lands around Dec 2026. After Dec 2026 it becomes harder to rule out RSI, though he still considers it unlikely.
More from AGI Musings
- Human taste can be gamed into optimized slop; preference is not a verifiable reward — torchcompiled · 2026-08-25
- Grove Research Founders Discuss Hugging Face Incident and Multi-Agent Dynamics — Miles_Brundage · 2026-08-25
- AI Should Exceed Benchmarks, Exploring New Frontiers of Human Potential — manosaie · 2026-08-25
- Viewpoint: Being Anti-AI Is Being Anti-Cancer and Longevity Research — justalexoki · 2026-08-25
- Writing Feedback Bottleneck Is Over: AI Changes Education — paulnovosad · 2026-08-25
- AI is the only major growth engine; without it, global outlook would be dramatically worse — iamtrask · 2026-08-25