Meta Muse Spark 1.1 Evaluation Released
ArtificialAnlys · x · 2026-07-11
Artificial Analysis reports that Meta's Muse Spark 1.1 scored 51 on the Intelligence Index, an 8-point increase over version 1.0, outperforming peers in both cost and token efficiency.
Key findings include:
- Significant gains in Scientific Reasoning, coding, and knowledge tasks, though it still trails stronger models in agentic knowledge work.
- Scored 45% on Humanity's Last Exam, closely trailing Claude Opus 4.8 (46%) and beating GPT-5.5 (44%) and Grok 4.5 (40%).
- Consumed only 94M output tokens to run the Intelligence Index, costing an estimated $0.26 per task—lower than GLM-5.2 and GPT-5.4.
- Other details: 1M context window, $1.25/$4.25 per million input/output tokens, and available directly via Meta's first-party API at launch.
More from Models
- Gemini 3.6 Flash appears live in Studio with $1.50 input pricing — ivan_bezdomny · 2026-07-21
- Artificial Analysis ranks Gemini 3.6 Flash at 50 on its updated intelligence index — Angaisb_ · 2026-07-21
- Google appears to have quietly shipped Gemini 3.6 Flash, with lower pricing and better agentic scores — xiaohu · 2026-07-21
- Google ships three more Gemini variants while 3.5 Pro slips again — Miserable-Archer-631 · 2026-07-21
- Google Quietly Launches Gemini 3.6 Flash: Cheaper, Stronger, and Agentic-Focused — OwariDa · 2026-07-21
- A user says 10–12 hours with Claude equals 3–4 hours with Grok Build — Daniel_Farinax · 2026-07-21