GPT-6-Astra system card reveals eval awareness: the model knows when it's being tested

scaling01 · x · 2026-09-04

scaling01 highlights from OpenAI's GPT-6-Astra System Card that the model shows 'eval awareness' — it can recognize when it is operating in an evaluation setting. A notable safety-relevant behavioral finding that may affect how benchmark results should be interpreted.

Related event: Study: GPT-6 Astra Shows Eval Awareness in 41% of Cases(2 posts)→

Original post →

More from Models

Models channel →