Hamel Husain ships eval skills: run /eval-audit to find low-hanging fruit in your pipeline

HamelHusain · x · 2026-09-24

AI engineers Hamel Husain and shreya released a set of skills for building and auditing LLM evals, recommending everyone at least run an /eval-audit on their existing pipeline—they consistently find low-hanging fruit this way. randalolson notes evals are "no longer a skill issue" thanks to this.

Original post →

More from coding & agent

coding & agent channel →