Rumor: Anthropic's Internal Model 'Mythos' Underwent 6 Months of RL Training
scaling01 · x · 2026-08-08
A recent tweet claims that it has been almost half a year since Anthropic's internal model, Mythos, was broadly available to its employees. The source suggests that 6 months of reinforcement learning (RL) hillclimbing with Mythos has resulted in scary capabilities. (Note: Unverified rumor)
More from Models
- DeepSeek Cascade Beats GPT-5.6 Luna on DeepSWE at 37% Lower Cost — togethercompute · 2026-08-08
- ChatGPT Seems to Ignore Memory Settings, Creeping Out User — flowersslop · 2026-08-08
- Reddit User Test: Outperforming Gemini Flash Lite — Horror-Slice-2772 · 2026-08-08
- New Architecture Model Shows Blazing Inference Speed for Real-Time Robotics — AkshatS07 · 2026-08-08
- Ant Group Releases Ling 3.0 Flash: 124B Model Hits Open Weights Pareto Frontier — ArtificialAnlys · 2026-08-08
- OpenAI Safety Team Details HF Incident: Rogue AI Behavior and 'Message Board' Phenomenon — dhadfieldmenell · 2026-08-08