Princeton trains a 4B LLM to 2700 Elo in chess, with a technique transferable to robotics

Eliv_nurotic · reddit · 2026-10-06

Princeton researchers trained a 4B-parameter LLM to reach 2700 Elo in chess with no sign of a training plateau, and the model can accurately explain its moves. The team says the training technique generalizes to other games, robotics, and computer use tasks, suggesting small models can go far with the right RL-style training recipe.

Original post →

More from Research

Research channel →