Gemini 3.8 held its Pareto frontier spot for just 3.5 hours before Muse Spark 1.3 undercut it
giffmana · x · 2026-09-03
- Artificial Analysis: Muse Spark 1.3 (xhigh) is the most cost-efficient model at its intelligence level — $0.55 per Intelligence Index task at Meta's unchanged $1.25/$4.25 per 1M token pricing
- Nearest rivals scoring 59+: Gemini 3.8 Flash (high, 59, $0.58), GPT-5.6 Sol (xhigh, 59, $0.63), GLM-5.3 (max, 60, $0.68); peers at 61 cost far more (Grok 4.6 $0.94, GPT-5.6 Sol max $0.95, Claude Opus 5 high $1.23)
- xeophon quips that Gemini 3.8 only held a Pareto frontier spot for 3.5 hours; Muse Spark 1.3 (max) excluded due to unannounced pricing
More from Models
- A Gemini Flash model reportedly tops the DeepSWE coding leaderboard — sunjiao123sun_ · 2026-09-03
- Hypothesis: the better LLMs get at coding, the worse their writing gets — kwangmoo_yi · 2026-09-03
- PINNACLE: Claude Fable 5.1 halves agent failure rate but costs $2.46 per correct answer — ryanshrout · 2026-09-03
- 5 more days of free MiniMax M3, M2.7, Music 3.0 and Speech 2.8 on GMI Cloud, plus $500 hackathon — MiniMax_AI · 2026-09-03
- Reddit user finds popular Qwen 27B finetune underperforms stock after thinking time cut — Brief-Effect9065 · 2026-09-03
- Fable 5.1 benchmarked against 8 rivals across coding, 3D, SVG and data-viz — arena · 2026-09-03