OpenAI previews Astra cybersecurity model reaching Critical threshold
OpenAI · x · 2026-09-02
OpenAI announced the upcoming release of Astra, a cybersecurity model that has reached the "Critical" threshold under its Preparedness Framework. The company emphasized its focus on making increasingly capable AI safe and broadly accessible. The post previews the evaluation methodology, advanced safeguards, and plans for continuous improvement.
Related event: OpenAI to release Astra, its first model with critical cyber capabilities(7 posts)→
More from Models
- Experts Question OpenAI Astra Eval Over Contamination and Metagaming Risks — ShakeelHashim · 2026-09-02
- Astra hits 100% success on ExploitBench refresh, reaching 'cyber-critical' threshold — infoxiao · 2026-09-02
- Anthropic Uses Activation Probes to Detect Cybersecurity Threats in Claude — nrehiew_ · 2026-09-02
- RWKV-7 G1j released: pure RNN architecture gets much better at agents and coding — jeremyphoward · 2026-09-02
- Fable 5.1 one-shots a working guitar VST plugin in 30 minutes — CtrlAltDwayne · 2026-09-02
- Fable 5.1 spontaneously solves 373-year-old cipher in 44 minutes — rickasaurus · 2026-09-02