Meta's Muse Spark 1.3: 38% fewer errors than Opus 5 at $0.85 per correct task

ryanshrout · x · 2026-09-04

Signal65's Ryan Shrout published new PINNACLE agentic benchmark results: Meta's Muse Spark 1.3 ranks second, with 38% fewer weighted errors than Claude Opus 5, and costs $0.85 per correct task versus $1.36 for Opus and $1.22 for GPT-5.6 Sol — knocking both hosted frontier models off the cost frontier.

Details:

This may be Meta's next "Llama moment."

Related event: Meta Muse Spark 1.3 Ranks Second on Agentic Benchmark at a Third of the Cost(2 posts)→

Original post →

More from Models

Models channel →