Not Anti-LLM-as-Judge: Regex Simply Can't Catch Semantic Nuances in Evals
xeophon · x · 2026-09-16
xeophon clarifies his earlier evals post wasn't anti-LLM-as-judge: catching semantic nuances (is "- <number>" a negative number or a bullet point?) is impossible with regex — exactly what LLM judges are for.
More from coding & agent
- Removed from org, 5 years of commits gone: dev can't train agent on own history — DanielLockyer · 2026-09-16
- Workshop Sep 19: explainable AI apps with Neo4j, GraphRAG, Cypher and LLM agents — camerongreen95 · 2026-09-16
- LLM bug hunt finds full-stack attack chain to brick a hardware device — matthew_d_green · 2026-09-16
- Claude Code spotted testing Sessions Hub: unified local and cloud session management — testingcatalog · 2026-09-16
- Rowboat Launches as Open-Source Multiplayer AI Assistant That Ships Code via Claude Code — ycombinator · 2026-09-16
- Dev builds FailEcho, a scanner that finds agent failures repeating across runs — EvenAd1183 · 2026-09-16