GPT-6 Astra tops DDD benchmark for multi-step retrosynthesis, nearing specialist models
CatAstro_Piyush · x · 2026-09-12
Vlad Aladin reports that in their DDD Benchmark, GPT-6 Astra outperformed every other frontier model tested on multi-step retrosynthesis, coming remarkably close to specialist-level synthesis planning models. The task predicts synthetic routes for small molecules — the same planning their Phase 2 molecule Rentosertib, which showed age-reversing effects across 6 aging clocks, required.
More from Models
- User: Claude Opus spun for 3 hours on a bug, Grok 4.6 fixed it in 15 minutes — Daniel_Farinax · 2026-09-12
- After a week of full-time use, Google's Astra has no opinions on anything — lucasmeijer · 2026-09-12
- LMArena analyzed 30,086 answer pairs: different LLMs share just 43.1% of ideas — arena · 2026-09-12
- Four Reported Tricks Behind "Dumbed-Down" Models: Routing, Juice Cuts, Truncated Reasoning, MTP — vista8 · 2026-09-12
- GPT Astra users fret over 'honeymoon window': compute shortage rumors spark performance anxiety — TooManyB1tches · 2026-09-12
- tldraw founder lists 10 bugs in ChatGPT's new sketch feature — manosaie · 2026-09-12