Stanford's Priced Guidance lower-bounds whether LLMs can forecast future research via paid hints

stanfordnlp · x · 2026-10-07

Percy Liang shares a new Stanford method: naively sampling whether an LLM can generate a future paper is intractable (astronomically small tail probability).

Priced Guidance instead has the model (Astra) play "20 Questions" with a guide LLM holding the paper, charging each hint by surprisal to precisely account for bits transferred — yielding a lower bound on P(LLM forecasts a future paper's idea). It doubles as a new eval for frontier models.

Related event: Stanford's Priced Guidance Measures LLM Novelty(3 posts)→

Original post →

More from Models

Models channel →