Frontier LLMs may judge arguments well, but not novelty
PlisSergey · x · 2026-07-26
The author argues that frontier LLMs are already better than humans at judging the quality of an argument.
But they should not be used to judge novelty or idea quality. In the author’s view, humans are already poor at that, but LLMs are even worse — and if novelty is delegated to models, the field risks “worshipping the phlogiston of the day.”
More from AGI Musings
- Programming Languages Becoming Computer-to-Computer Communication, Researcher Notes — eptwts · 2026-07-26
- Jensen Huang says agents are the new software and AI will create more jobs than it kills — garrytan · 2026-07-26
- Survey: Many Willing to Adopt Human Enhancement Tech, But Want Regulation — Chris_Armstrong · 2026-07-26
- Why AI's Economic Impact Seems Small: Lessons from the Industrial Revolution — sebkrier · 2026-07-26
- Investor Warns: Heavily Funding Free Open-Source LLMs is a High-Risk Strategy — ghadfield · 2026-07-26
- AI is Unreliable for Research? Pair Generation with Robust Verification — joshgans · 2026-07-26