OpenAI Safety Team Warns Astra Model Reaches 'Critical' Cyber Capability

eyishazyer · x · 2026-08-08

OpenAI's safety team reported that internal evals of the upcoming Astra model show exceptional performance in agentic coding and cyber tasks, reaching a potential 'Critical' capability tier. This means Astra could autonomously build working zero-day exploits or execute end-to-end cyberattacks. The company is tightening security and pausing internal projects that don't meet these elevated safety bars.

Related event: OpenAI Slows Astra Development Over Cyber Risk(26 posts)→

Original post →

More from Models

Models channel →