GPT-6 System Card's log-scale graph obscures Astra's CoT controllability jump, Reddit user argues

thegamebegins25 · reddit · 2026-09-04

Reddit user thegamebegins25 argues a chart in OpenAI's GPT-6 Astra System Card is "bordering on deceptive":

This matters because CoT monitoring is currently one of our strongest safety monitors; if models can control their CoT far more often, monitor effectiveness drops sharply. Source: OpenAI's deploymentsafety page, figure 28.

Original post →

More from Models

Models channel →