Reasoning-depth estimate casts doubt on Grok 6.1 looping gains

scaling01 · x · 2026-09-30

scaling01 shares a discussion around a "true reasoning depth" estimate applied to Grok models. The results are contradictory: either looping isn't very effective, 6.1 Sol has a smaller base model than 6 Sol (which he doubts), or 6.1 Sol isn't looping at all. His conclusion: the metric is likely unreliable and a single loop doesn't double reasoning depth.

Related event: Benchmark sleuths suspect Grok 6.1 of answer-reusing loops(3 posts)→

Original post →

More from Models

Models channel →