Autorubric at COLM 2026: a unifying framework for rubric-based LLM evaluation

deliprao · x · 2026-10-09

Delip Rao's team is presenting Autorubric at COLM 2026, a unifying framework for rubric-based LLM evaluation on non-verifiable tasks, aimed at practitioners doing evals and reward modeling.

The idea is to make open-ended tasks gradeable via rubrics, supporting both rubric optimization and downstream objective optimization. A companion cookbook and runnable Python examples followed in the thread.

Related event: Autorubric: A Unified Framework for Rubric-Based LLM Evaluation Unveiled at COLM 2026(4 posts)→

Original post →

More from Research

Research channel →