Rumor: Anthropic's Internal Model 'Mythos' Underwent 6 Months of RL Training

scaling01 · x · 2026-08-08

A recent tweet claims that it has been almost half a year since Anthropic's internal model, Mythos, was broadly available to its employees. The source suggests that 6 months of reinforcement learning (RL) hillclimbing with Mythos has resulted in scary capabilities. (Note: Unverified rumor)

Original post →

More from Models

Models channel →