Report: OpenAI's Astra may use technique that destroys CoT monitorability
sjgadler · x · 2026-09-02
Citing a report by The Information, DavidKasten highlights that OpenAI may have utilized a breakthrough in 'neuralese' for Astra which could destroy chain-of-thought monitorability. While sources claim OpenAI is currently 'limiting the use of the technique,' concerns arise regarding the ambiguity of 'limiting' and the practical enforcement of monitoring, especially given OpenAI's previous explicit discouragement of such practices.
Related event: OpenAI's Astra Reportedly Uses Recurrent Depth for Silent Latent Reasoning(29 posts)→
More from Models
- GLM-5.3 Hits 310 tok/s, Coding Performance Competes with Opus — Yuchenj_UW · 2026-09-02
- User cancels Claude Max over confusing rate limits and new restrictions — robleclerc · 2026-09-02
- Claude Fable 5.1 crushes hard coding benchmarks, outpaces Chinese models — minchoi · 2026-09-02
- Fable 5.1 recreates an Airbus H145 helicopter in Three.js from a simple prompt — minchoi · 2026-09-02
- Gary Marcus: OpenAI's new technique could destroy chain-of-thought monitorability — GaryMarcus · 2026-09-02
- Anthropic releases official prompting guide for Claude 5.1 — ethanCaballero · 2026-09-02