LlamaIndex's Jerry Liu and Snorkel AI on why evals and RL environments remain unsolved

ajratner · x · 2026-09-30

Jerry Liu hosted a dinner with Snorkel AI's Vincent Chen on evals and RL environments: data/RL-env companies are booming, benchmarks get shredded each release, fairness attribution (input vs harness vs reward) is hard, and raw enterprise data must be "developed" into usable evals and training environments via new engineering paradigms involving domain experts.

Original post →

More from coding & agent

coding & agent channel →