Thread argues scaled RL and large distillation favor Anthropic in the current race
rickasaurus · x · 2026-07-25
The post reacts to a thread arguing that scaled RL plus large-scale distillation is powerful, with Anthropic seen as ahead while OpenAI may still catch up with bigger models. The underlying claim is that post-training expertise can materially shift model rankings.
More from Models
- Claude Opus 5 can misread a document despite knowing the underlying facts — teortaxesTex · 2026-07-25
- Early Opus 5 feedback says Claude’s writing is now “4o-level slop,” despite stronger intelligence — jdjohnson · 2026-07-25
- Users say Anthropic’s new model feels faster and stronger than Fable — emax · 2026-07-25
- Claude Opus 5 scores 30.2% on ARC-AGI-3 public demo environments — GregKamradt · 2026-07-25
- Release blog teaser shows a near-tie on FrontierCode agentic coding benchmark — hardmaru · 2026-07-25
- User Accuses Anthropic of Gaming ARC-AGI-3 by Training Specifically on Benchmark Patterns — VraserX · 2026-07-25