Muse Spark 1.1 Shows Major Evaluation Gains

alexandr_wang · x · 2026-07-09

This post outlines the core capabilities of Muse Spark 1.1: it approaches the performance of GPT-5.5 and Opus 4.8 in multiple agentic evaluations, showing significant improvements over its predecessor in tool use, computer use, long-context coding, and visual reasoning.

The post also notes that the model supports a 1M context window, features actual computer use capabilities, and offers a public API, emphasizing its overall competence and release details.

Related event: Meta Launches Muse Spark 1.1: A Low-Cost, High-Performance Agentic Model(133 posts)→

Original post →

More from Models

Models channel →