LangChain team: evals are the #1 blocker keeping companies from becoming AI-native

hwchase17 · x · 2026-10-06

LangChain's Brace Sproul says evals are the biggest blocker in most agent development pipelines, with changes coming to LangSmith "in a few short weeks."

An accompanying post lays out the core definitions: an eval is a grader applied to a trace; a grader is anything that scores agent performance (from a simple code check to an LLM-as-judge); a trace is the full record of one agent run—input, every tool call (e.g. searchorders, issuerefund), and output.

The takeaway: hand-tweaking prompts and testing a few inputs only gets you so far—evals are what get agents deployment-ready and keep them working in production, and they're the top hurdle to companies becoming truly AI-native.

Related event: LangChain Team Highlights Evals as Top Barrier to Enterprise AI Adoption(2 posts)→

Original post →

More from coding & agent

coding & agent channel →