Replaying 100 human group discussions with LLM agents: 98.7% consensus, mostly on wrong answers

alex_verem · x · 2026-10-07

A Waseda University researcher's arXiv paper (2609.20543) replayed 100 real human Wason-card group discussions with belief-anchored LLM agent groups.

Key findings:

Conclusion: simulated consensus does not track collective accuracy, and belief-anchored agent groups are biased estimators of real human deliberation — a warning for social-simulation research.

Related event: Waseda Study Finds AI Focus Groups Agree Often — But Often Wrong Together(2 posts)→

Original post →

More from AGI Musings

AGI Musings channel →