HalluWorld: Controlled LLM Hallucination Benchmark Finds Perception Solved, Simulation Still Hard

Jeande_d · x · 2026-09-29

HalluWorld is a controlled hallucination benchmark built on fully-specified reference worlds (grid worlds, chess, terminals), accepted to NeurIPS 2026 Evaluations & Datasets. A model hallucinates when it makes an observable claim false in the reference world.

Key findings:

The paper argues hallucination is not one capability, and existing benchmarks are too fragmented to compare mitigations across settings.

Original post →

More from Research

Research channel →