OpenAI's GPT-6 Astra hits Critical cybersecurity threshold, first model to do so

moyix · x · 2026-09-04

OpenAI released the GPT-6 Astra system card: its most capable deployed model and the first to reach the Critical cybersecurity level under its Preparedness Framework—able to find unknown flaws and develop exploits across hardened systems with minimal human guidance. Third-party evals say Astra performed long-horizon vuln research and autonomously exploited multiple 0days. Mitigations include stricter isolation, checkpoint encryption, full-trajectory monitoring including CoT, blocking alignment evals, and greatly improved jailbreak robustness versus GPT-5.6 Sol.

Related event: OpenAI Releases GPT-6 Astra System Card, First Model Rated Critical for Cybersecurity(11 posts)→

Original post →

More from Models

Models channel →