AA-Omniscience Eval: Sol and V4 Lead in Knowledge but Are 'Bullshit Machines'

teortaxesTex · x · 2026-07-21

A user shared interesting findings from the AA-Omniscience evaluation:

These results prompt a discussion on the actual effects of scaling on model hallucinations and knowledge retention.

Related event: AA-Omniscience Chart Compares Accuracy and Hallucination Rates of 28 Models(2 posts)→

Original post →

More from Models

Models channel →