Top AI Models Get About a Quarter of Factual Questions Wrong Without Search

Data from Artificial Analysis' AA-Omniscience benchmark, covering 6,000 factual questions across 42 domains, shows top AI models get roughly a quarter of answers wrong without search, highlighting the need for verification of important facts.

2026-09-30 ~ 2026-09-30 · 2 related posts