OpenAI's Astra Cybersecurity Model Reaches 'Critical' Threshold
alexcovo_eth · x · 2026-09-02
OpenAI announced the upcoming release of Astra, a cybersecurity model that has reached the "Critical" threshold under its Preparedness Framework.
The Critical threshold is met if the model can either:
- Identify and develop functional zero-day exploits of all severity levels in many hardened real-world critical systems without human intervention.
- Devise and execute end-to-end novel strategies for cyberattacks against hardened targets given only a high-level goal.
OpenAI previewed its evaluation methods, the safeguards advanced alongside capabilities, and plans for continuous improvement.
More from Models
- Dev: Anthropic 5.1 is a massive leap over previous model — doodlestein · 2026-09-02
- Fable 5.1 test fails to meet expectations — Angaisb_ · 2026-09-02
- Claude Fable 5.1 generates a full Mario Kart game with a single prompt — ezshine · 2026-09-02
- Gemini 3.7 Flash Demonstrates Agentic Video Understanding — otarU · 2026-09-02
- OpenAI's Astra achieves 100% success rate on ExploitBench vulnerability tests — JiaweiLiu_ · 2026-09-02
- Testing Claude Fable 5.1: Max mode produces the best pelican SVG for $3.30 — Simon Willison · 2026-09-02