Meta Muse Spark 1.1 Evaluation Released
ArtificialAnlys · x · 2026-07-11
Artificial Analysis reports that Meta's Muse Spark 1.1 scored 51 on the Intelligence Index, an 8-point increase over version 1.0, outperforming peers in both cost and token efficiency.
Key findings include:
- Significant gains in Scientific Reasoning, coding, and knowledge tasks, though it still trails stronger models in agentic knowledge work.
- Scored 45% on Humanity's Last Exam, closely trailing Claude Opus 4.8 (46%) and beating GPT-5.5 (44%) and Grok 4.5 (40%).
- Consumed only 94M output tokens to run the Intelligence Index, costing an estimated $0.26 per task—lower than GLM-5.2 and GPT-5.4.
- Other details: 1M context window, $1.25/$4.25 per million input/output tokens, and available directly via Meta's first-party API at launch.
More from Models
- BullshitBench update: GPT-6-Astra beats all prior OpenAI models but still trails Anthropic — scaling01 · 2026-09-11
- Astra Scores 83% on GauntletBench, First Computer-Use Agent to Beat Human Baseline — ducha_aiki · 2026-09-11
- Kimi K2.8 Preview rolls out: near-K3 coding performance, 1M context for all tiers — teortaxesTex · 2026-09-11
- Looking for a classifier of software engineering task shapes to pick models per task — StewartalsopIII · 2026-09-11
- DeepSeek V4 Pro API to continue after Sept 2026, billing unchanged — teortaxesTex · 2026-09-11
- DeepSeek V4.1 Flash Hits 98% of GPT-6 Astra's Score at 1.4% of the Cost in Third-Party Benchmark — ayushtweetshere · 2026-09-11