GPT-4 stays accurate on unpublished social science studies and open-weight models also perform well

RobbWiller · x · 2026-07-23

Model comparison

The paper reports that GPT-4’s predictions remained strongly correlated with observed effects even when the experiments were not public by the training cutoff or were still unpublished in June 2025.

Robustness

Takeaway

The result is not just a memorization artifact: performance stayed high on unpublished or later-published studies as well.

Related event: Nature Study: GPT-4 Can Predict Social Science Experiment Results(15 posts)→

Original post →

More from Research

Research channel →