Report: ByteDance Pre-training Massive Model Nearing 10T Parameters
zijing_wu · x · 2026-08-07
According to Zijing Wu, ByteDance is targeting a mega AI model nearing 10 trillion (10T) parameters in the early stage of pre-training, aiming to compete with Anthropic's estimated 8T parameter Mythos model.
Additionally, multiple Chinese labs are training models around 5T parameters (e.g., Fable est. 5T), while Seed has been pursuing a no-distillation approach for over a year.
Related event: Report: ByteDance Preps 5 to 10 Trillion Parameter AI Model(4 posts)→
More from Models
- Users Complain About Claude Opus 5's Poor Performance, Anticipate Quick Replacement — KlausCodes · 2026-08-07
- Ling 3.0 Tiny Supports Native Function Calling with Only 1.3B Active Parameters — Danare_113 · 2026-08-07
- Latent Space Briefing: Meta Wins Olympiad Golds, OpenAI Unifies Models and MCP Ecosystem — Latent Space · 2026-08-07
- Elon Musk Announces Grok Build v1.0: Free CLI Coding Agent Powered by Grok 4.5 — elonmusk · 2026-08-07
- OpenAI's Logan Kilpatrick Teases 'Great New Models' Are in the Oven — emollick · 2026-08-07
- Professor Finds AI Elaborately Cheating to Win at Nethack — emollick · 2026-08-07