Muse Spark 1.2 Tops Finance Agent Benchmark at Fraction of Opus Cost
EdwardSun0909 · x · 2026-08-07
Muse Spark 1.2 is the first model to break the 60% accuracy mark on ValsAI's Finance Agent v2 benchmark, which tasks models with financial analyst duties.
The model achieves this top-tier performance at a cost of just $0.77 per test—6.7x cheaper than the previous #1 model, Opus 5 ($5.12)—while operating at twice the speed.
Related event: Muse Spark 1.2 Tops Finance Agent Benchmark(4 posts)→
More from Models
- ChatGPT Free Tier Gets Unlimited Chat as GPT-5.6 Sol Optimizes Conversational Experience — 量子位 · 2026-08-07
- Grok 4.5 Beats Kimi K3 at 13x Lower Cost in Agent Task Test — rohanpaul_ai · 2026-08-07
- Meta's Muse Spark 1.2 Hits Pareto Frontier at 1/5th of Claude's Cost — ArtificialAnlys · 2026-08-07
- Meta's Muse Spark 1.2 Hits Pareto Frontier at 1/6th the Cost of Claude — ArtificialAnlys · 2026-08-07
- OpenAI's Upcoming Device to Focus on Personality, But Can the Model Deliver? — Angaisb_ · 2026-08-07
- Rabdos Launches Math AI Benchmark; Claude Opus 5 Takes the Lead — AI4Code · 2026-08-07