GPT-6 System Card's log-scale graph obscures Astra's CoT controllability jump, Reddit user argues
thegamebegins25 · reddit · 2026-09-04
A Reddit user argues OpenAI's log-scale chart in the GPT-6 Astra System Card obscures a major change: Astra maintains near-100% chain-of-thought controllability up to a few hundred tokens, versus 10% for models earlier this summer — meaning the model can defeat CoT monitors far more often, undermining one of our strongest safety monitoring techniques.
More from Models
- GPT-6 Pro limited to 50 messages/week on the $100 plan, users complain — TrustedEssentials · 2026-09-04
- Unverified leak claims OpenAI ARC-AGI-3 score jumped from 7.8% to 98.6% — miilesus · 2026-09-04
- GPT-6 ties with Grok 4.6 and Muse Spark 1.3 on benchmark — ns123abc · 2026-09-04
- OpenAI's GPT-6 Astra hits Critical cybersecurity threshold, first model to do so — moyix · 2026-09-04
- GPT-6 Astra crushes ARC-AGI-3: 62.7% standard, 99.9% with adapter harness — Hesamation · 2026-09-04
- Astra Beats GPT Pro at Proof-Checking, a First for Non-Pro Models — joshgans · 2026-09-04