DeepSeek Performance Debate: "10T Model" Rumors Dismissed
Users pushed back on community claims that DeepSeek flopped and relied on a secret "10T parameter" or "Mythos teacher" model, citing a 16B model scoring 80% on ARC-AGI-2 and noting there is no evidence for the rumors.
2026-08-15 ~ 2026-08-15 · 2 related posts
- Debate: DeepSeek Performance and Skepticism About "10T Models" — teortaxesTex · 2026-08-15
- DeepSeek New Model Sparks Debate: 16B Hits 80% on ARC-AGI-2, Distillation Rumors Questioned — teortaxesTex · 2026-08-15