New Preprint with Tetlock: How RL Scoring Rules Reshape LLM Forecasting Behavior

simonguozirui · x · 2026-09-03

A new preprint from Lightning Rod AI with forecasting experts Philip Tetlock and Ville Satopää post-trains 5 versions of the same LLM, varying only the scoring rule used as the RL reward. Similar aggregate scores emerge, but very different BIN profiles. Key points:

The takeaway: choosing the scoring rule is a critical design decision in training LLM forecasters, trading off accuracy and error types.

Original post →

More from Research

Research channel →