Only Muse Spark 1.3 and Fable 5.1 sit on the coding Pareto frontier
jyangballin · x · 2026-09-11
- EdwardSun0909 shared a coding benchmark chart showing only Muse Spark 1.3 and Fable 5.1 remain on the performance-cost Pareto frontier.
- leerob confirmed Grok 4.6's lower score is accurate as measured, and teased that Grok 4.7 is coming soon.
More from Models
- OpenAI pauses new $200 ChatGPT Pro sign-ups as heavy users burn through quota — MickeySteamboat · 2026-09-11
- Genspark launches Gen-1 Slides, a work model priced at 1/17th of Opus 5 — Scobleizer · 2026-09-11
- OpenAI's Rumored Internal Model 'Bel' Could Be the Most Hyped Release Ever — imadade · 2026-09-11
- DeepSeek-V4.1-Flash rolls out to Ollama cloud Pro subscribers after Max and Team debut — ollama · 2026-09-11
- GPT-6 Astra called a computer-use model built for knowledge work like accounting — mckbrando · 2026-09-11
- Sakana AI launches Fugu Max and Fugu Ultra v2, matching elite models at 2-6x lower cost — SakanaAILabs · 2026-09-11