Oxford-led NeurIPS paper: frontier LRMs mirror human rule discovery and brain activity

sreejan_kumar · x · 2026-09-25

A NeurIPS-accepted study from Oxford, Columbia, NYU, MIT and Harvard had 32 fMRI-scanned humans and frontier LRMs play VGDL grid games with no rules given. Top LRMs matched human learning curves, discovering rules in similar steps; model hidden states predicted human BOLD responses across cortical and subcortical regions — the first demonstration that LRM representations align with human brain activity during active learning. Ablations attribute alignment to in-context representations of game-state sequences.

Original post →

More from Research

Research channel →