Cheap LLMs all middling at narrow academic task, not trivial after all

jon_mellon · x · 2026-09-19

RexDouglass expected a narrow task to work well, but jonmellon notes most cheap models they tried were middling at detecting null findings in academic abstracts, so it's not a trivial task.

Related event: Researchers Find Budget Open Models Struggle to Detect Null Findings, Newer Models Show Promise(9 posts)→

Original post →

More from Models

Models channel →