Benchmark Score Gap Attributed to Safeguards, Not Smarts
eyishazyer · x · 2026-09-02
Comparative data reveals that Fable 5.1 and Mythos 5.1 share the same model weights. The score gap on Terminal-Bench 4.0 (55.8 vs 60.9) is entirely due to differences in safeguards, not increased intelligence.
More from Safety
- Study: Slanted framing of real images outpaces deepfakes in multimodal misinformation — IAugenstein · 2026-09-02
- Azure OpenAI Flaw May Have Exposed SharePoint Data to Unauthorized Users — emmanuelvivier · 2026-09-02
- CoT Monitoring Fragility: Models Struggle to Verbalize Tool-Returned Cues — PMinervini · 2026-09-02
- Indian Copyright Office registers AI-generated work, attributing authorship to creator — technollama · 2026-09-02
- Anthropic Paper Reveals Risks of Deception in CoT Monitoring — raphaelmilliere · 2026-09-02
- Spiking NN Lib BindsNET Compromised in DPRK Supply Chain Attack — cyb3rops · 2026-09-02