Text-only reasoning solves Rubik's Cube: GPT-5.6 Sol clears CubeBench's 0% barrier in 333 moves

Willing_Plate_5417 · reddit · 2026-09-12

CubeBench previously reported a 0.00% pass rate for leading LLMs on long-horizon Rubik's Cube tasks testing spatial reasoning and state tracking. A Reddit user re-ran the experiment with a frontier model under strict constraints—no Python, no external solver, no hidden solution, just text-based reasoning against a cube engine.

GPT-5.6 Sol solved a freshly scrambled 3×3: 750.5 seconds, 333 moves, 61 interactions, 0 rejected moves. The solution was inefficient with self-correction, but it recognized OLL/PLL cases along the way. The takeaway: '0% success' timestamps a generation of models, not a permanent architectural boundary.

Original post →

More from Models

Models channel →