Western Open-Weight Models Are Catching Up to Chinese Ones
scaling01 · x · 2026-07-16
The author believes people should stop asking "how far behind are Chinese open-weight models compared to Western ones," and instead ask "how far behind are Western open-weight models compared to Chinese ones."
The post quotes a lengthy review of a new open-source/open-weight model: it features an MoE architecture with 975B total parameters / 41B activated parameters, is trained on 45 trillion tokens, and supports text, image, and audio reasoning. However, its benchmark scores aren't particularly impressive, feeling like another Kimi-K2.6, and to some, it seems rushed out before subsequent versions are fully ready.
More from Models
- China’s AI arms race is increasingly defined by chips, data centers, and open models — BenBajarin · 2026-07-22
- Sam Altman is headed to Washington to brief Congress on OpenAI’s GPT-6 line — inductionheads · 2026-07-22
- Benchmark chart pits GPT-5.6 Luna, Grok 4.5 and Gemini 3.6 Flash on price and scores — iruletheworldmo · 2026-07-22
- Claim says Kimi was distilled from Fable, sparking a model-attribution jab — cephaloform · 2026-07-22
- Gemini 3.6 Flash is now available in Antigravity and chat — MartianOnJupiter · 2026-07-22
- Gary Marcus says LLM math skills are like knowing only a car’s engine size — GaryMarcus · 2026-07-22