burny pushes back on 'LLMs just predict likely text': RLHF, CoT and tools changed that

burny_tech · x · 2026-10-01

Leaker burnytech quotes tech journalist Taylor Lorenz and argues her framing is outdated: the 'LLMs just output the most likely text sequences' mental model is 2+ years old. Modern models' output distributions are heavily shaped by RL from rewards, plus CoT, tool use and RAG. Lorenz agrees frontier models are great but hopes error rates improve.

Original post →

More from Models

Models channel →