Muse Spark 1.1 Beats GLM-5.2 in Coding Eval
The Decoder · rss · 2026-07-11
The Decoder reports that Meta's Muse Spark 1.1 scored 51 on the Artificial Analysis Intelligence Index, marking an 8-point improvement in just three months.
Key data points from the report include:
- For coding tasks, Muse Spark 1.1 scored 71.3, slightly outperforming GLM-5.2.
- The cost per task is approximately $0.26, making it cheaper than its counterparts.
- The hallucination rate dropped from 73% to 38%.
The core takeaway is that the model has made solid progress in coding ability, cost efficiency, and hallucination control.
More from Models
- Users say GPT-5.6 Ultra feels like extra token burn with little visible gain — CtrlAltDwayne · 2026-07-21
- Early Gemini 3.6 Flash outputs look fast but weak on frontend and spatial reasoning — max_paperclips · 2026-07-21
- Anthropic removes Fable’s access deadline, but users say it was nerfed — oykun · 2026-07-21
- Kimi K3 retakes first place on DesignArena’s frontend web app benchmark — rohanpaul_ai · 2026-07-21
- Last Week in AI roundup covers Claude Sonnet 5, LongCat 2.0, and new agent benchmarks — Last Week in AI · 2026-07-21
- Rumor claims GPT-6 could arrive in August — iruletheworldmo · 2026-07-21