Google Paper: Autonomous AI Research Hallucinates 90% Without Checks
rohanpaul_ai · x · 2026-09-01
A new Google paper reveals the risks of autonomous AI research. Without reliability modules, severe result hallucinations appeared in 90% of Agent Laboratory papers and 46% of Co-Scientist papers. When Co-Scientist cross-referenced claims with execution logs, the hallucination rate dropped to 4% and complete data fabrication to 0%.
More from Safety
- Would OpenAI survive a near-miss liability regime after the HF hack? — dfrsrchtwts · 2026-09-01
- Report: OpenAI and Anthropic Paused RL Training — tszzl · 2026-09-01
- Scholars propose using LLMs for pre-review in academic peer review — anderssandberg · 2026-09-01
- Paper defines cognition-induced risks in Agentic AI systems — 机器之心 · 2026-09-01
- Agents can't verify people: data enrichment APIs are failing — Dry_Steak30 · 2026-09-01
- Are model mental motions equivalent to unconscious reward expectations? — FioraStarlight · 2026-09-01