LLM Forecasting Study: Activations More Reliable Than Text

burny_tech · x · 2026-07-11

Goodfire released a paper on LLM forecasting titled "What LLM Forecasters Know but Don’t Say".

The paper points out that while LLM forecasters often appear confident, their calibration is unstable, and their chain-of-thought may not reveal why predictions change. The research found that internal model activations reflect the true state better than the output text.

By using small probes to read these activations, the study aims to:

The authors argue this offers a cheaper and more faithful approach to auditing LLM predictions.

Original post →

More from Research

Research channel →