Frontier LLMs may judge arguments well, but not novelty

PlisSergey · x · 2026-07-26

The author argues that frontier LLMs are already better than humans at judging the quality of an argument.

But they should not be used to judge novelty or idea quality. In the author’s view, humans are already poor at that, but LLMs are even worse — and if novelty is delegated to models, the field risks “worshipping the phlogiston of the day.”

Original post →

More from AGI Musings

AGI Musings channel →