DeepSeek New Model Sparks Debate: 16B Hits 80% on ARC-AGI-2, Distillation Rumors Questioned
teortaxesTex · x · 2026-08-15
On X, user @teortaxesTex responds to rumors about DeepSeek, stating DeepSeek did not 'shit the bed' and questioning evidence for '10T models' or 'Mythos teacher'. He notes a 16B model achieving 80% on ARC-AGI-2, arguing that 1-2 evals are insufficient to prove Mythos-Preview > Mythos. He also observes inference fleet expansion, suggesting cost reduction as evidence.
Related event: DeepSeek Performance Debate: "10T Model" Rumors Dismissed(2 posts)→
More from Models
- DeepSeek Developing Flash Variants to Match 3T Model Coding Performance — bindureddy · 2026-08-15
- Qwen3.8-27B performance sparks buzz, netizens joke about Google's reaction — max_paperclips · 2026-08-15
- Developer quantizes AI9Stars G9v3-39A5B to GGUF and creates llama.cpp fork for support — linuxid10t · 2026-08-15
- Gemini 3.5 Pro checkpoint renamed to Gemini 3.7 Flash High, sparking speculation — Rare_Bunch4348 · 2026-08-15
- Perplexity releases Agent API and web search benchmarks — AravSrinivas · 2026-08-15
- Debate: DeepSeek Performance and Skepticism About "10T Models" — teortaxesTex · 2026-08-15