OpenAI's GPT-6 Astra hits Critical cybersecurity threshold, first model to do so
moyix · x · 2026-09-04
OpenAI released the GPT-6 Astra system card: its most capable deployed model and the first to reach the Critical cybersecurity level under its Preparedness Framework—able to find unknown flaws and develop exploits across hardened systems with minimal human guidance. Third-party evals say Astra performed long-horizon vuln research and autonomously exploited multiple 0days. Mitigations include stricter isolation, checkpoint encryption, full-trajectory monitoring including CoT, blocking alignment evals, and greatly improved jailbreak robustness versus GPT-5.6 Sol.
More from Models
- Leak: GPT-6 Astra's Training Incubation Ran Early May To Late July — scaling01 · 2026-09-04
- GPT-6 Astra Won't Charge Extra Usage In Codex Until 272k Tokens — pvncher · 2026-09-04
- Reddit user: GPT-6 Astra is surprisingly good at circuit design and chip architecture — Christs_Elite · 2026-09-04
- Mirai's uzu engine brings speculative decoding to Apple M5, hitting 105 tok/s on Qwen3.6 27B — TheMoonMidas · 2026-09-04
- Best local models for 12GB of VRAM: Gemma-4-12B remains the pick — GlennCameronjr · 2026-09-04
- Claude Fable 5.1 Launches, Early Users Say It One-Shots the Best Websites of Any Model — repligate · 2026-09-04