Dev claims Meta's Muse Spark 1.3 only looks good on benchmarks, 'severely gamed'
bindureddy · x · 2026-09-03
Developer Bindu Reddy claims Meta's Muse Spark 1.3 "only looks good on benchmarks," rating it in real-world use as merely a "GPT 5.6 Terra class model" and asserting that "the benchmarks are being gamed severely." Unverified third-party commentary worth tracking in the ongoing debate over benchmark scores versus real-world performance.
Related event: Founder Claims Meta Muse Spark 1.3 Benchmarks Are Misleading(2 posts)→
More from Models
- Claude is down for many users amid widespread outage — Polymarket · 2026-09-03
- Claude Mythos 5.1, Fable 5.1 and Opus 5 hit elevated errors, Anthropic investigating — ClaudeAI-mod-bot · 2026-09-03
- Two-Author Model Tech Report Praised as Dense: Pretraining to Downstream — antoine_chaffin · 2026-09-03
- Leaked hints suggest the next release will be v2.5 — koltregaskes · 2026-09-03
- MBZUAI releases K2 Horizon: six fully open models from 0.9B to 375B with training code and data — kimmonismus · 2026-09-03
- OpenAI's naming chaos: GPT-4.5 and o1 were each once slated to be GPT-5 — flowersslop · 2026-09-03