Fable 5.1 tops vision+logic benchmark near the top; scores 78 on private logic test vs prior high of 61
Afinetheorem · x · 2026-09-02
An independent tester's quick take on Fable 5.1: it still writes in that very Claude way; it's the first non-Gemini model to rank right near the top on a vision+logic benchmark; and it scored 78 on a private logic benchmark written for another lab, well above the previous high of 61 (Sol 5.6 Pro).
Related event: Fable 5.1 Sets Record on Private Vision-Logic Benchmark(3 posts)→
More from Models
- World Labs releases Atlas: A pixel-perfect camera controlled world model — viksit · 2026-09-02
- Fable 5.1 Science Benchmark: Autonomous success rate doubles to 53.6% — johnseach · 2026-09-02
- Anthropic releases Claude 5.1 with 25% lower costs and zero data retention — eugeneyan · 2026-09-02
- OpenAI previews Astra cybersecurity model reaching Critical threshold — OpenAI · 2026-09-02
- Observation: AI agents become succinct in voice mode, adapting to human listeners — joshwhiton · 2026-09-02
- Anthropic Accused of Retroactively Adding Safeguards to Older Opus Models — LordCoice · 2026-09-02