Dietterich: LLMs Go Beyond Next-Token Prediction

ML researcher Thomas Dietterich pushed back on claims that LLMs merely predict the next word, arguing that SFT and RL have moved models toward predicting correct multi-token sequences, and that pretraining itself forces abstract generalization beyond parroting.

2026-09-25 ~ 2026-09-25 · 2 related posts