Mitchell clarifies paper's LLM definition, sparking debate on post-trained models

mmitchell_ai · x · 2026-09-28

In a debate with Blanche Minerva and Melanie Mitchell, mmitchellai clarifies that his paper's use of "LLM" specifically meant "pretrained" and "unsupervised" LMs. Minerva pushes back with a thought experiment: a transformer pretrained on text but then taught tool-calling tokens, command line use, CoT reasoning/planning, and multi-turn conversation — is it still an LLM under that definition? The thread probes whether the classic terminology survives modern post-training pipelines.

Related event: Melanie Mitchell's 'These Are Not LLMs' Sparks Fresh Stochastic Parrots Debate(31 posts)→

Original post →

More from AGI Musings

AGI Musings channel →