Mathematician: latest models Sol 5.6 and Fable 5.1 can't solve any of my real problems
burny_tech · x · 2026-09-07
Pushing back on the "math is dead" takes, a researcher notes how many ordinary math problems AI still fails. Testing Sol 5.6 and Fable 5.1 on problems he has genuinely been thinking about, neither model solved a single one—a counterpoint to headlines of models acing benchmark math.
More from Models
- Gary Marcus asks for a full timeline of OpenAI's hacking incident disclosures — GaryMarcus · 2026-09-07
- GPT-6 Astra generates animated black hole scene with custom WebGL/GLSL shaders — omarsar0 · 2026-09-07
- 146 model variants tested on creative writing: Opus, Fable, Kimi K3 top the board — zainhas · 2026-09-07
- Claude's Writing Tics: Personified Subjects and Arguing by Negation, Dissected — Sauers_ · 2026-09-07
- GPT-6 Astra on Low Beats GPT-5.6 Sol on High, Devs Advise Lower Effort — reach_vb · 2026-09-07
- After 115K videos in prod, engineer shares Gemini video-understanding gotchas and hacks — TheMoonMidas · 2026-09-07