Muse Spark 1.1 Evaluation Score Jumps 8 Points
ArtificialAnlys · x · 2026-07-11
Artificial Analysis revealed that Meta's Muse Spark 1.1 improved by 8 points over version 1.0 on the Intelligence Index, reaching 51 points.
The growth is primarily concentrated in:
- Scientific reasoning, programming, and knowledge.
- Coding Index: 59 → 71 (+12)
- SciCode: 52% → 58% (+6)
- Humanity's Last Exam: 40% → 45% (+5)
- AA-Omniscience: 4 → 18 (+14)
- GDPval-AA v2: 1144 → 1376 Elo (+232)
The post also notes that it is now on par with models like GLM-5.2, GPT-5.4, and GPT-5.6 Luna, though it still trails behind the highest-scoring tier of models.
More from Models
- Astra Scores 83% on GauntletBench, First Computer-Use Agent to Beat Human Baseline — ducha_aiki · 2026-09-11
- Kimi K2.8 Preview rolls out: near-K3 coding performance, 1M context for all tiers — teortaxesTex · 2026-09-11
- Looking for a classifier of software engineering task shapes to pick models per task — StewartalsopIII · 2026-09-11
- DeepSeek V4 Pro API to continue after Sept 2026, billing unchanged — teortaxesTex · 2026-09-11
- DeepSeek V4.1 Flash Hits 98% of GPT-6 Astra's Score at 1.4% of the Cost in Third-Party Benchmark — ayushtweetshere · 2026-09-11
- TheZvi Polls: Has Your Coding Model Choice Changed Since Fable 5.1 and Astra? — TheZvi · 2026-09-11