Anthropic Report Reveals Internal Model Exceeding Mythos 5

In its second 186-page Risk Report, Anthropic disclosed that a model with the internal codename "Model 2" slightly outperforms Claude Mythos 5 overall and is already heavily used for coding, data generation, research, and agentic tasks—though the company made clear there are "no plans to release it externally." The model scored 62.8% on the CoBench v2 test versus 50.3% for Mythos 5, a 12.5-percentage-point gap, but it still falls short of the 85% target for fully automated researcher replacement.

Confirmed

Unconfirmed

2026-08-15 ~ 2026-08-15 · 11 related posts

Primary sources

1 near-duplicate retellings: kimmonismus