OpenAI Pauses High-Risk Astra but Ships GPT-5.6-Cyber
eyishazyer · x · 2026-08-11
A user highlights that OpenAI made opposite calls on two highly capable cybersecurity models within the same week.
Key Events:
- Pausing Astra: On August 7, OpenAI paused the Astra model internally because it might hit their top cyber-risk tier. This model could independently find and chain zero-days against hardened targets.
- Shipping GPT-5.6-Cyber: On August 10, they shipped GPT-5.6-Cyber, which sits one risk rung lower. It answers 95% of exploit-development queries and has even found a real zero-day in Chrome's V8 engine (patched under a live CVE).
The author notes this shows OpenAI's safety framework working as intended, drawing a hard line at Critical while treating lower tiers as usable with access controls. However, it also implies that the capability "ceiling" companies are comfortable shipping is steadily creeping upward.
Related event: OpenAI Faces Internal Dispute Over AI Model Safety Ratings(2 posts)→
More from Models
- River AI Raises $1.1B to Build Custom Agents and LLMs via API — Teknium · 2026-08-11
- UnslothAI Confirms Its Acceleration Tools Work Well with Apple's MLX Framework — danielhanchen · 2026-08-11
- How to Strip Claude's Text Watermark? Users Test Translation Workarounds — churchkey · 2026-08-11
- Anthropic Criticized for AI Watermark Strategy That Could Drive Users Away — Brian821 · 2026-08-11
- Daybreak Blue in Codex Clarified: Not GPT-5.6 Cyber, But Security-Tailored GPT-5.6 Sol — Angaisb_ · 2026-08-11
- Anthropic to Embed Imperceptible Watermarks in Claude Text for EU Compliance — lilyraynyc · 2026-08-11