OpenAI Safety Team Warns Astra Model Reaches 'Critical' Cyber Capability
eyishazyer · x · 2026-08-08
OpenAI's safety team reported that internal evals of the upcoming Astra model show exceptional performance in agentic coding and cyber tasks, reaching a potential 'Critical' capability tier. This means Astra could autonomously build working zero-day exploits or execute end-to-end cyberattacks. The company is tightening security and pausing internal projects that don't meet these elevated safety bars.
Related event: OpenAI Slows Astra Development Over Cyber Risk(26 posts)→
More from Models
- Grok's Image Model Impresses with Exceptional Background Text Rendering — XFreeze · 2026-08-08
- Why are frontier AI models suddenly speaking in 'caveman speak'? — zainhas · 2026-08-08
- xAI Releases Imagine Image 2.0, Ranking Just Behind OpenAI in Arena Benchmarks — The Decoder · 2026-08-08
- Reddit User Comparison: Claude Still the Best Overall, GPT and Kimi Close Behind — pbad1 · 2026-08-08
- llama.cpp Adds Support for Longcat-Flash Model, Open for Testing — pmttyji · 2026-08-08
- OpenAI Launches Continuous Voice Mode as Astra Stuns in Math — eyishazyer · 2026-08-08