Anthropic discloses Claude now leads 26% of its R&D, with 30,000 agents running daily
新智元 · wechat · 2026-09-18
Anthropic has for the first time published three internal metrics on the pace of AI development. Using Epoch's six-level automation scale, Claude now autonomously "leads" about 26% of core R&D work, over 90% of R&D involves deep human-AI collaboration, and no subtask has reached full L5 autonomy. Roughly 30,000 R&D agents run concurrently on internal platforms, governed by two layers of oversight: pre-execution real-time scanning (over 1B decisions tracked in August, 0.002% blocked) plus weekly offline review funneling 100,000 suspicious logs down to 50 top-priority incidents for human experts. Agents get independent identity labels and communicate on a transparent public message bus to prevent collusion.
The compute ledger is equally candid: only 6% of core R&D compute goes to safety research (12% for AI-led work), which Anthropic calls a deliberately conservative accounting. The disclosure aims to set an industry transparency baseline, challenging other frontier labs (read: OpenAI) to publish equivalent metrics verifiable by third parties such as METR.
Related event: Anthropic Discloses Internal Metrics: Claude Now Drives 26% of Its R&D(41 posts)→
More from AGI Musings
- François Fleuret's alien joke: after 3.8B years, carbon-based design is about to switch — francoisfleuret · 2026-09-21
- Cheap human-level intelligence could build a Dyson sphere in a decade — EigenGender · 2026-09-21
- "Let em rip": X users debate how far current models are from real danger — tszzl · 2026-09-21
- Frontier open-source models have been out for ages and nothing happened, doomers told — tszzl · 2026-09-21
- Debate: regimes don't need x-risk-level AI to harm their residents — amplifiedamp · 2026-09-21
- Economist's "messy jobs" theory: two traits that protect work from AI displacement — ksprdk · 2026-09-21