HCI papers increasingly use LLM judges while obfuscating it, researcher warns
IanArawjo · x · 2026-09-03
Researcher Ian Arawjo reports finding more and more papers at HCI venues that use LLM judges, with a worrying trend: the writing obfuscates that LLM judges were used at all — sometimes it's hard to even tell what "IRR" (inter-rater reliability) refers to. The observation raises academic transparency and reviewing-norm concerns.
Related event: HCI Papers Quietly Use LLM Judges, Obscuring the Fact(2 posts)→
More from Research
- SimLoss Enables Single-Pass Fine-Grained Image Captioning at Multi-Stage Quality — Suryaansh Jain · 2026-09-03
- Tenstorrent and AI & Inc launch JapanFold: free inference for open-source drug discovery models — DavidBennett__ · 2026-09-03
- Until Labs Scales Cryoprotectant Search to 250,000 Molecules With AI — NirantK · 2026-09-03
- X Debate: Is Chain-of-Thought Prompting a Form of Parameter Reuse? — aryaman2020 · 2026-09-03
- Constraining agents with LL(1) grammar + structured diagnostics: what it fixes and what slips through — Upstairs-Special-925 · 2026-09-03
- Nora optimizer keeps Muon's matrix structure benefits without the full cost — burkov · 2026-09-03