OpenAI Rates GPT-6 Astra 'Critical' for Cyber Capabilities, Powered by 100% ExploitBench Scores

fouadmatin · x · 2026-09-04

OpenAI's Fouad Matin details the safety evaluation behind deeming GPT-6 Astra to have Critical cyber capabilities: 100% of the determination came from ExploitBench, where the model passed even at low reasoning effort.

The team is now evaluating against more recent real-world vulnerabilities and building harder benchmarks, calling it 'the AGI era for cybersecurity.'

A rare official rating of a frontier lab's flagship model as Critical-level in offensive cyber capability.

Related event: OpenAI's Astra aces ExploitBench with 100% exploit rate, rated first Critical-level cybersecurity model(11 posts)→

Original post →

More from Models

Models channel →