Reasoning-depth estimate casts doubt on Grok 6.1 looping gains
scaling01 · x · 2026-09-30
scaling01 shares a discussion around a "true reasoning depth" estimate applied to Grok models. The results are contradictory: either looping isn't very effective, 6.1 Sol has a smaller base model than 6 Sol (which he doubts), or 6.1 Sol isn't looping at all. His conclusion: the metric is likely unreliable and a single loop doesn't double reasoning depth.
Related event: Benchmark sleuths suspect Grok 6.1 of answer-reusing loops(3 posts)→
More from Models
- Claim: Opus 5.5 one-shot a full music video entirely in code, no video model — repligate · 2026-09-30
- Reddit user reports receiving $62,500 in surprise ChatGPT API credits — KeyBaker5 · 2026-09-30
- dots local agent tasks: why do different models burn quota at wildly different rates? — lxfater · 2026-09-30
- User cancels ChatGPT for Claude, citing stingy limits and quota-reset tactics — ezshine · 2026-09-30
- Claude has gotten a lot better at creating infographics over the past few months — tom_doerr · 2026-09-30
- DepthBench compares 10 architectures to find which residual tweaks actually buy computational depth — SonglinYang4 · 2026-09-30