Nature paper finds GPT-4 can predict social science experiment results with r=0.85
JeremyNguyenPhD · x · 2026-07-24
A Nature paper reports that LLMs can predict outcomes of social science experiments with strong accuracy.
- The authors built an archive of 70 preregistered, nationally representative U.S. survey experiments, covering 469 treatment effects and 119,330 participants.
- Using GPT-4 to simulate responses from representative American samples, they inferred treatment effects by comparing simulated answers across conditions.
- The predicted effects were strongly correlated with observed results, reaching r = 0.85, and performed similarly to pooled human forecasters.
- The correlation remained high even for studies not published before the model’s training cutoff, and for predictions from prominent open-weight models.
- A notable caveat: the models systematically overestimated effect sizes.
- In a secondary archive of 15 megastudies with 606 effects, correlations were lower but still comparable to pooled expert forecasters.
- The authors argue LLMs may help with pilot testing, intervention selection, and spotting effects that need replication, while also raising concerns about bias and misuse.
Related event: Nature Study: GPT-4 Can Predict Social Science Experiment Results(18 posts)→
More from Research
- Bug Hunt Bench author: leaderboard noise is about 2-3 points — PawelHuryn · 2026-09-11
- Bug Hunt Bench ranks frontier coding models on 105 planted real-repo bugs — PawelHuryn · 2026-09-11
- PNAS paper shows a tiny billiard-ball system is a universal computer — undecidability lives in two dimensions — eigensteve · 2026-09-11
- New paper: Absolute pose estimation from affine cues and gravity direction — ducha_aiki · 2026-09-11
- LoMa Paper Ships REALLY HardPairs Dataset, Accepted at ECCV 2026 — ducha_aiki · 2026-09-11
- Johns Hopkins Launches Full-Stack Hands-on Robot Learning Class with SO-101 Arm Kits — _krishna_murthy · 2026-09-11