Curbing Hallucinations is Harder: Rooted in Pre-training

teortaxesTex · x · 2026-07-07

The author finds a certain model's progress in reducing hallucinations particularly impressive.

The argument is that positive cognitive abilities like reasoning are sparse and can be pushed almost indefinitely using RL. However, hallucinations seem more closely tied to foundational knowledge and pre-training (unless one "cheats" by artificially inflating the refusal rate to bypass the issue).

Original post →

More from Models

Models channel →