Fable 5.1 more than doubles score on agentic scientific workflows, 24.7% to 52.6%
haider1 · x · 2026-09-02
haider1 reports that Fable 5.1 jumped from 24.7% to 52.6% on agentic scientific workflows — more than 2x better than its predecessor. He ties this to RL prioritization, noting rumors that OpenAI's unreleased models are exceptionally strong at math, suggesting labs can push fields one by one by choosing what to optimize.
More from Models
- OpenAI and Others Quietly Using Loop Transformers That Hide Their Thinking — harris_edouard · 2026-09-02
- Anthropic launches browser-based C2PA checker to detect Claude-made images, video and audio — jedisct1 · 2026-09-02
- Early Hands-On: Fable 5.1 Looks Promising So Far — james_mtc · 2026-09-02
- AISLE finds 6 curl CVEs days after OpenAI and Anthropic security systems reported zero — stanislavfort · 2026-09-02
- OpenAI reportedly passed on the GPT-6 name for Astra, saving the 6-worthy jump for Bel — Angaisb_ · 2026-09-02
- Early user review: hy4 preview goes down the right path but reaches wrong conclusions — xeophon · 2026-09-02