Nature paper finds GPT-4 can predict social science experiment results with r=0.85
JeremyNguyenPhD · x · 2026-07-24
A Nature paper reports that LLMs can predict outcomes of social science experiments with strong accuracy.
- The authors built an archive of 70 preregistered, nationally representative U.S. survey experiments, covering 469 treatment effects and 119,330 participants.
- Using GPT-4 to simulate responses from representative American samples, they inferred treatment effects by comparing simulated answers across conditions.
- The predicted effects were strongly correlated with observed results, reaching r = 0.85, and performed similarly to pooled human forecasters.
- The correlation remained high even for studies not published before the model’s training cutoff, and for predictions from prominent open-weight models.
- A notable caveat: the models systematically overestimated effect sizes.
- In a secondary archive of 15 megastudies with 606 effects, correlations were lower but still comparable to pooled expert forecasters.
- The authors argue LLMs may help with pilot testing, intervention selection, and spotting effects that need replication, while also raising concerns about bias and misuse.
Related event: Nature Study: GPT-4 Can Predict Social Science Experiment Results(18 posts)→
More from Research
- Causal-only attention for non-generative tasks is wasteful, argues HF engineer — antoine_chaffin · 2026-09-11
- Catholic University of Chile researcher: scaling AI feedback is key to sustainable medical education — julianvarascom · 2026-09-11
- Nature paper images cellular activity across all organs, revealing body-wide circuits — arjunrajlab · 2026-09-11
- SignNet 1M Dataset Released for Sign Language Research — ducha_aiki · 2026-09-11
- ECCV26 Oral: Flow Matching Enables Single-Stage Multi-View Point Cloud Registration — ducha_aiki · 2026-09-11
- InFlux++ Method Released — ducha_aiki · 2026-09-11