Google's AI Overview Flips Answer Based on Single arXiv Preprint

sayashk · x · 2026-08-05

After releasing a paper on whether AI agents can do open-ended research, the authors found that Google's AI Overview reversed its answer to their main research question from a confident "Yes" to a confident "No" within hours.

The authors highlight the double-edged nature of this behavior. While updating responses based on new evidence is desirable, it is concerning that a single, unreviewed arXiv paper was enough to completely flip the authoritative summary. This demonstrates how easily these responses could be manipulated with adversarial intent. A human expert would weigh one new paper against the entire existing literature more carefully before updating their take.

Original post →

More from Safety

Safety channel →