Anthropic discloses Claude now leads 26% of its R&D, with 30,000 agents running daily

新智元 · wechat · 2026-09-18

Anthropic has for the first time published three internal metrics on the pace of AI development. Using Epoch's six-level automation scale, Claude now autonomously "leads" about 26% of core R&D work, over 90% of R&D involves deep human-AI collaboration, and no subtask has reached full L5 autonomy. Roughly 30,000 R&D agents run concurrently on internal platforms, governed by two layers of oversight: pre-execution real-time scanning (over 1B decisions tracked in August, 0.002% blocked) plus weekly offline review funneling 100,000 suspicious logs down to 50 top-priority incidents for human experts. Agents get independent identity labels and communicate on a transparent public message bus to prevent collusion.

The compute ledger is equally candid: only 6% of core R&D compute goes to safety research (12% for AI-led work), which Anthropic calls a deliberately conservative accounting. The disclosure aims to set an industry transparency baseline, challenging other frontier labs (read: OpenAI) to publish equivalent metrics verifiable by third parties such as METR.

Related event: Anthropic Discloses Internal Metrics: Claude Now Drives 26% of Its R&D(41 posts)→

Original post →

More from AGI Musings

AGI Musings channel →