DRACO benchmark still puts fusion systems ahead of single frontier models
iamtrask · x · 2026-07-22
OpenMined reran an OpenRouter-style experiment on DRACO, a deep-research benchmark, and found that fusion models still dominate.
- The top 5 systems are all fusions, and the best single frontier model only placed 6th.
- The post says this matches observations from four startups and Los Alamos National Lab.
- The chart shows fusion combinations such as Fable + GPT-5.5, Opus + GPT-5.5 + DeepSeek, and DeepSeek + Kimi + GPT-5.5 outperforming solo models by a wide margin.
The author frames this as evidence that combining models remains the best path on deep-research tasks, and that more benchmarks are coming.
Related event: DRACO Benchmark Shows Routing Systems Outperform Single Frontier Models(2 posts)→
More from Research
- 3D ResNet Paper Crosses 3,000 Citations Eight Years After CVPR 2018 — HirokatuKataoka · 2026-09-11
- Sample selection and ordering matter a lot in LLM training: DataFlex makes data scheduling dynamic — Puzzleheaded_Box2842 · 2026-09-11
- Jeff Heaton's Intro to the Math of Neural Networks eBook Is Free to Download — blaizedsouza · 2026-09-11
- Mathematician Daniel Litt Launches Problem Repo to Track Human vs AI Progress: 15 Problems, 1 Solved — littmath · 2026-09-11
- Open ECDSA.fail challenge uses AI agents to shrink Shor's-algorithm quantum circuits for Bitcoin keys — StefanoGogioso · 2026-09-11
- Alex Townsend posts 200 open problems in numerical linear algebra for humans and AI agents — IgorCarron · 2026-09-11