GPT-6 Astra sets new SOTA on HealthBench Professional at half the cost of GPT-5.6 Sol
thekaransinghal · x · 2026-09-04
OpenAI's Karan Singhal details GPT-6 Astra's medical performance:
- Astra sets a new state of the art on HealthBench Professional, a benchmark of real clinician tasks covering care consultations, medical documentation, and research
- Even at its lowest reasoning effort, Astra surpasses GPT-5.6 Sol's best score at roughly half the cost
- The team says the evaluation was intentionally designed to be challenging for their own models
This is Astra's first public vertical-domain result, with cost-efficiency in medical workloads as the headline claim.
Related event: OpenAI's GPT-6 Astra Sets New Record on HealthBench Professional(3 posts)→
More from Models
- OpenAI launches GPT-6 Astra, claiming it can do anything you do on a computer — kagigz · 2026-09-04
- Databricks evals: GPT-6 Astra claims SOTA on OfficeQA Pro benchmarks, cheaper per task — downingARK · 2026-09-04
- Mathematician tests GPT-6 Astra: live Lean proof verification while writing arguments — teortaxesTex · 2026-09-04
- Tavus Launches Sparrow-2, Claiming #1 in End-of-Turn Detection and Interruption Handling — ycombinator · 2026-09-04
- ARC-AGI-3: 10x reasoning tokens cuts total cost from $48k to $26k vs medium — i_dg23 · 2026-09-04
- Researchers flag data contamination concerns in benchmark behind Astra's time-horizon score — dfrsrchtwts · 2026-09-04