FULL STORY

Meta Launches Muse Spark 1.3 as Reviews Roll In

Meta released Muse Spark 1.3, its fourth iteration in five months, scoring 62 on intelligence benchmarks. Developers soon tested its self-improvement capabilities, with Meta AI chief Alexandr Wang amplifying the results.

2026-09-03 ~ 2026-09-03 · 4 episodes · 51 posts

Episode 1 · Meta releases Muse Spark 1.3, closing the gap with frontier models (2026-09-03, 44 posts)

Meta released its proprietary reasoning model Muse Spark 1.3 on September 3, the fourth iteration in five months. Artificial Analysis has published its eval page: the max version scores 62 on the Intelligence Index and the xhigh version 61, a continued sharp climb from 53 for 1.1 and 57 for 1.2, far above the category median of 17 and ranking near the top among all 636 models. Reddit users conclude Meta is slowly closing the gap with top models like GPT-5.6 Sol.

Confirmed

  • Artificial Analysis Intelligence Index: xhigh 61, max 62—on par with GPT-5.6 Sol (max)
  • Proprietary reasoning model with a 1M tokens context window
  • Pricing stays at $1.25/$4.25 per million tokens; the xhigh version costs just $0.55 per Intelligence Index task, the most cost-efficient model at its intelligence level
  • Tau3-Bench Banking: max scores 52%, a new first place on that benchmark; xhigh hits 47%, comparable to Claude Fable 5.1 (max) and peers
  • Gains over 1.2 concentrate in agentic evals, with a marked jump on GDPval-AA v2

Why it matters

  • Four releases in five months plus consecutive Intelligence Index jumps show Meta rapidly closing the gap with frontier models
  • Lowest task cost at its intelligence tier is directly attractive for usage-based Agent applications
  • Agent-focused evals (GDPval-AA v2, Tau3-Bench Banking) drove most of the gains, showing the iteration is aimed squarely at agentic capability

24 more related posts →

Episode 2 · Meta Releases Muse Spark 1.3 as Its Strongest Model (2026-09-03, 2 posts)

Meta has released Muse Spark 1.3, its strongest model yet, with Alexandr Wang claiming it beats GPT-5.6 Sol on coding. Polymarket traders remain skeptical, giving Meta only a 12% chance of holding the top AI model by year-end.

Episode 3 · Meta's Muse Spark Runs 2-Hour Self-Improvement Loops for Under $1, Stunning Scale CEO (2026-09-03, 3 posts)

A user benchmark found Meta's Muse Spark 1.3 Ultra finished a task in about a minute while Fable 5.1 took 70 minutes and $13, and ran 20 self-improvement rounds with 60+ agent operations for under $1 over two hours, stunning Scale AI CEO Alexandr Wang.

Episode 4 · Meta Chief AI Officer Praises muse spark 1.3 Sub-Agent Demo (2026-09-03, 2 posts)

Meta's Chief AI Officer Alexandr Wang shared a developer's hands-on test of muse spark 1.3, calling it "pretty cool." The developer deployed parallel sub-agents that completed the task in about an hour.