GPT-4 matches 2,659 human forecasters and improves when averaged with them
RobbWiller · x · 2026-07-23
Human forecast comparison
GPT-4 matched the pooled forecasts of 2,659 people.
Complementarity
Human and GPT-4 predictions were not redundant: averaging them improved accuracy beyond either one alone, reaching r = .89.
What this adds
The figure suggests the model is not only accurate on its own, but also provides information that can complement crowd forecasts.
Related event: Nature Study: GPT-4 Can Predict Social Science Experiment Results(15 posts)→
More from Research
- Robotics paper says VLA and world models are not enough for grounded supervision — hbouammar · 2026-07-23
- AI could compress decades of biomedical research into days, says Derya Unutmaz — DeryaTR_ · 2026-07-23
- OpenAI and Apollo show models may optimize graders, not user intent — rohanpaul_ai · 2026-07-23
- AI Autonomously Disproves Decades-Old Math Conjectures: The Singularity's Opening Phase — imjustnewatai · 2026-07-23
- Applied Math Dominates AI, But Why Does Gradient Descent Actually Work? — fkasummer · 2026-07-23
- Cursor’s Composer 2.5 looks much worse at reasoning than its Kimi base model — gleech · 2026-07-23