Muse Spark 1.1 Ranks High in Coding Agent Eval
ArtificialAnlys · x · 2026-07-11
Muse Spark 1.1 (xhigh) scored 69 on the Artificial Analysis Coding Agent Index using the Opencode harness. The post provides the following comparisons:
- Slightly lower than GPT-5.5 (medium), which scored 71 in Codex.
- Higher than Claude Opus 4.8 (medium), which scored 67 in Claude Code.
- At roughly $1.4 per task, it is highly cost-effective compared to other frontier coding agents, though the trade-off is a longer processing time per task.
More from coding & agent
- alphaXiv open-sources OpenResearch to run parallel research agents with any model — alphaXiv · 2026-09-11
- MathModelAgent gains traction: auto-solves math modeling and writes a submission-ready paper — jihe520 · 2026-09-11
- DeskcommCRM: open-source AI sales CRM with native agents and WhatsApp hits 1k stars — melgarafael · 2026-09-11
- hyperresearch: agent-driven knowledge base that turns web research into a searchable wiki — jordan-gibbs · 2026-09-11
- Forter's 13 lessons from its agent sprint: skip custom RAG, lean on mature enterprise search — bibryam · 2026-09-11
- Two real 'company brains' opened up live: Gorgias' in-house Cortex vs Slite — femke_plantinga · 2026-09-11