Meta's Muse Spark 1.3 tops coding benchmarks vs GPT 5.6 and Opus 5, open weights coming
armand_ruiz · x · 2026-09-03
- Meta shipped Muse Spark 1.3, claiming its biggest jump yet on coding and agentic work at near-commodity pricing, live in Muse Code and the API.
- Third-party benchmarks show strong coding results: DeepSWE v1.1 at 75.4 beats GPT 5.6 Sol (73.0) and Opus 5 (74.0); codebase understanding (SWEAtlas) 59.4 clears both by 6+ points; Terminal-Bench 2.1 at 88.8 ties GPT and beats Opus's 86.
- Meta also teased an upcoming model (watermelon codename) and open-weight Muse Spark releases.
Related event: Meta launches Muse Spark 1.3, matching Fable 5 at a fraction of the price(53 posts)→
More from Models
- GPT-6 Astra and Astra Aeon spotted in Codex; Aeon tipped as long-horizon agent model — VraserX · 2026-09-03
- Ant Group Open-Sources Finance-Enhanced Model Ling-3.0-flash-Fin with 124B Parameters — niacolhealth · 2026-09-03
- Should 'hard scientific problems solved' be the new LLM benchmark? — Dr_Singularity · 2026-09-03
- Ben Thompson on Gemini: 'OK, it's now or never' — kieranklaassen · 2026-09-03
- User Reports ChatGPT Randomly Outputs 'GOD IS COMING' and Won't Stop — jaypro1005 · 2026-09-03
- Ling-3.0-flash-Fin Weights Released: 124B Params, 5.1B Active, 256K Context — Bestlife73 · 2026-09-03