Muse Spark 1.1 Scores 69 in Coding Benchmark
ArtificialAnlys · x · 2026-07-12
According to Artificial Analysis, Muse Spark 1.1 (xhigh) scored 69 on their Coding Agent Index, performing near the frontier under the Opencode harness with strong cost efficiency. For context, it scored slightly lower than GPT-5.5 (medium) at 71, but higher than Claude Opus 4.8 (medium) at 67. The post also notes that its cost per task is roughly $1.4, making it one of the cheaper frontier coding agents, though this comes with certain capability trade-offs.
Related event: Meta Muse Spark 1.1 Shines Across Multiple Benchmarks(13 posts)→
More from coding & agent
- Devin adds e2b sandboxes for remote agent execution — badphilosopher · 2026-07-22
- Hermes Agent Refactoring Proposal: Decoupling via Event Bus and Monorepo Slicing — Promptmethus · 2026-07-22
- ty now reads Pydantic config keywords and field metadata — charliermarsh · 2026-07-22
- Pensar Launches AI Security Agent to Autonomously Discover and Patch 0-Days — andriy_mulyar · 2026-07-22
- ty adds first-class Pydantic support, including strict and lax field handling — charliermarsh · 2026-07-22
- Google launches Gemini 3.5 Flash Cyber for CodeMender, with limited access for governments — GoogleAI · 2026-07-22