Report: OpenAI Paused Astra for Cyber Risk, Then Shipped GPT-5.6-Cyber Days Later
SimplyAnnisa · x · 2026-08-11
A recent tweet highlights a controversy over OpenAI's internal safety risk ratings. On August 7, OpenAI reportedly paused Project Astra internally because a model could hit their top cyber-risk tier, capable of independently finding and chaining zero-day vulnerabilities.
However, just three days later on August 10, OpenAI shipped GPT-5.6-Cyber, which was rated "High" (one rung down). This model answers 95% of the exact kind of sensitive vulnerability questions, raising questions about the consistency of safety thresholds.
Related event: OpenAI Faces Internal Dispute Over AI Model Safety Ratings(2 posts)→
More from Models
- Where Do Quantized Local LLMs Break? Reddit Users Share Experiences — d77chong · 2026-08-12
- Unreleased Anthropic Model Makes Surprising Progress on the Riemann Hypothesis — TechCrunch AI · 2026-08-12
- Frontier Models' Biggest Bottleneck is 'Lack of Self-Esteem' in Math Research — coherence · 2026-08-12
- Pokee-Isaac 28B Beats Meta and Qwen Rivals in Sub-30B Agent Benchmarks — Kyrannio · 2026-08-12
- Post-Trained NVIDIA Nemotron Beats Claude Opus in Legal Agent Tasks — ctnzr · 2026-08-12
- New LLM Jailbreak Trick: Just Tell It to "Be Smarter Than Grok" — npinto · 2026-08-12