Mixing 8% flawed reasoning traces lifts a chess LLM to 1300 Elo

amasad · x · 2026-07-29

A chess LLM training experiment found that pure board→move supervision hit strong diminishing returns. An accidental branch that tried to “think before you move” produced hallucinated moves, but mixing just about 8% of those reasoning traces into the training mix beat pure move data at the same compute budget.

Related event: LLMs Still Struggle with Chess Despite Pre-training(3 posts)→

Original post →

More from Research

Research channel →