Have Chinese Models Caught Up With the US Frontier?
scaling01 · x · 2026-07-19
This deep dive attempts to answer whether Chinese AI models have caught up with the US frontier. The author notes that while the Artificial Intelligence Index isn't ideal for directly measuring true model strength, the article later uses it to predict the ECI of Kimi-K3, making the approach somewhat contradictory.
The author compares two metrics: Artificial Intelligence Index and ECI. The conclusion is that AAI is less discriminative than ECI regarding model strength, and a fitted log-linear regression reveals that "Chinese models are overestimated on AAI." Specifically, only 36.7% of Chinese models score above the regression line, compared to 57.5% of US models.
Based on this regression line, the author estimates a predicted ECI of 158.33 for Kimi-K3, with an 80% confidence interval. The article's focus isn't a single leaderboard, but rather an exploration of the US-China model capability gap and metric biases using multiple public indicators.
Related event: Has China Caught Up to US Frontier AI? The Debate Reignites(7 posts)→
More from Models
- BullshitBench update: GPT-6-Astra beats all prior OpenAI models but still trails Anthropic — scaling01 · 2026-09-11
- Astra Scores 83% on GauntletBench, First Computer-Use Agent to Beat Human Baseline — ducha_aiki · 2026-09-11
- Kimi K2.8 Preview rolls out: near-K3 coding performance, 1M context for all tiers — teortaxesTex · 2026-09-11
- Looking for a classifier of software engineering task shapes to pick models per task — StewartalsopIII · 2026-09-11
- DeepSeek V4 Pro API to continue after Sept 2026, billing unchanged — teortaxesTex · 2026-09-11
- DeepSeek V4.1 Flash Hits 98% of GPT-6 Astra's Score at 1.4% of the Cost in Third-Party Benchmark — ayushtweetshere · 2026-09-11