OpenAI releases GPT-6 Astra: first model to hit Critical cyber capability level

PeterHndrsn · x · 2026-09-04

OpenAI has released GPT-6 Astra, its most capable broadly deployed model and the first to reach the Critical level of cybersecurity capability under its Preparedness Framework—able, with the right tools, to find unknown flaws and develop new exploits across well-protected systems without step-by-step human guidance.

Key safety measures:

Korbak notes Astra is more aligned but less monitorable, attributing this to an intelligence jump. Observers warn that governments using such models for offensive cyber operations could cause large off-target effects.

Related event: GPT-6 Astra Launch: Capability Leap Marred by Benchmarking Dispute and Declining Monitorability(62 posts)→

Original post →

More from Models

Models channel →