New paper: LLM outputs are far less diverse than their training data, and temperature tuning won't fix it

rohanpaul_ai · x · 2026-09-04

The paper "Do LLMs Capture the Diversity in their Training Data?" compares model continuations against training continuations for the same prefixes:

Takeaway: LLMs learn a narrower answer distribution than their training data—a structural phenomenon beyond sampling strategy.

Original post →

More from Research

Research channel →