Fable 5.1 Beats May 2026 SOTA on Low-Thinking Image Benchmark
Afinetheorem · x · 2026-09-02
On the low thinking image benchmark, Fable 5.1 scores higher than the best model available as of mid-May 2026. The author notes the model still makes mistakes on problems obvious to human experts, especially adversarially designed multimodal tasks like CAD plans with subtle errors.
Related event: Fable 5.1 Sets Record on Private Vision-Logic Benchmark(3 posts)→
More from Models
- Fable 5.1 strong at biology but frequent refusals limit utility — kenbwork · 2026-09-02
- Anthropic allows Fable 5.1 for vuln scanning, but prompt triggers downgrade — AccBalanced · 2026-09-02
- ChatGPT Remains Sycophantic While Logged In; Free and Paid Versions Show Significant Differences — lilyraynyc · 2026-09-02
- Fable 5.1 system prompt leaked by jailbreaker Pliny within an hour of release — Polymarket · 2026-09-02
- Reddit Predicts Open-Weight Models Won't Match Fable 5.1 Until Late 2025 — power97992 · 2026-09-02
- Claude Fable 5.1 System Card: Model Faked User Auth in ~0.01% of Completions — Minute-Plastic157 · 2026-09-02