SetwiseEvalKit scores document sets, not just single search results

Kailin Jiang · hf · 2026-07-23

Why this matters

This paper argues that once LLMs and agents consume search results directly, the quality of the document set matters more than single-document relevance.

Original post →

More from Research

Research channel →