Claude Fable 5.1 scores 86.6% on SimpleBench, beating human baseline
Profanion · reddit · 2026-09-03
SimpleBench results shared on Reddit:
- Claude Fable 5.1: 86.6%
- Human baseline: 83.7%
- Gemini 3.8 Flash: 82.4%
- Claude Fable: 81.9%
- Muse Spark 1.3: 81.8%
Claude Fable 5.1 tops the human baseline.
More from Models
- OpenEvidence Launches Medical AI Model Family, Darwin Scores First-Ever 100% on MedQA — benxneo · 2026-09-03
- New data: open-weight models are already in production worldwide, and cost isn't the top reason — perilli · 2026-09-03
- "Why are all the major providers down?" — multi-provider AI outage confuses users — basedjensen · 2026-09-03
- ChatGPT down today amid apparent multi-provider AI outage — Abhishekcur · 2026-09-03
- Local model users get the last laugh as major cloud AI providers go down — zeeg · 2026-09-03
- Open Athena Starts Training Marin: 535B-Parameter MoE on 18T Tokens, Fully Open — TheTuringPost · 2026-09-03