Muse Spark 1.3 Tops DeepSWE at 75.4%, Beating GPT-5.6 Sol and Fable 5

kimmonismus · x · 2026-09-03

Muse Spark 1.3 has reportedly taken first place on DeepSWE v1.1 with 75.4%, ahead of GPT-5.6 Sol and Fable 5, drawing surprise reactions.

The SoTA score reinforces Meta's claim of a major coding and agentic leap, shaking up the coding-benchmark leaderboard.

Related event: Muse Spark 1.3 Tops DeepSWE Benchmark at 75.4%(2 posts)→

Original post →

More from Models

Models channel →