Melanie Mitchell: Post-Training Like RLHF Makes Models No Longer Statistical Language Models
MelMitchell1 · x · 2026-09-27
Santa Fe Institute's Melanie Mitchell weighs into a terminology debate: people now use "LLM" loosely for any large AI system built on a pre-trained LM, which is not the sense meant in the Stochastic Parrots paper — that's her only point.
She adds that any post-training moving the model away from the general statistics of language (e.g., RLHF or RL for tool use) makes it no longer a statistical model of language in the sense of modeling the structure of human language.
More from Models
- One persona tweak made ChatGPT say 'goblins' 4,000% more — caught on Reddit before OpenAI noticed — victor_explore · 2026-09-27
- First PhD paper accepted at NeurIPS 2026: Sparse layers key to scaling looped LMs — burny_tech · 2026-09-27
- MiMo-V2.6 listing hints at 5 models: 1T, 311B and 9B visible, two more unknown — jacek2023 · 2026-09-27
- antirez: Good programmers failing with GPT 6 Astra points to a different skill set — antirez · 2026-09-27
- Opus 5.5 and GPT-6 shipped 101 minutes apart, both cheaper — airesearch12 · 2026-09-27
- Hands-on: Opus 5.5 nails frontend consistency; Astra still wins reasoning — haider1 · 2026-09-27