Those Model 'Thinking' Phrases Aren't Real Reasoning Traces

rao2z · x · 2026-10-06

Researcher rao2z clarifies that the comforting phrases frontier models flash while users wait are not actual chain-of-thought traces. Since o1, labs hide real traces from paying users (who still pay for those tokens), showing generated "commentary" instead — the notable exception being DeepSeek-R1, which exposes full, often incoherent, page-long reasoning.

Related event: Researcher Warns Models' Flashed 'Thinking' Phrases Are Not Real Chain-of-Thought(2 posts)→

Original post →

More from Models

Models channel →