Engineer's eval: cross-encoder rerankers lift 30% of cases but wreck the retriever in another 30%

bo_wangbo · x · 2026-10-08

Engineer bowangbo shares first-hand eval results on cross-encoder rerankers: on his eval sets, they improved 30% of cases, did nothing in 40%, and ruined the first-stage retriever in another 30%, leaving him little confidence in cross-encoders for OOD tasks. The thread also asks whether late-interaction models count as cross-encoders.

Related event: Debate: Are Cross-Encoder Rerankers Dead?(4 posts)→

Original post →

More from Models

Models channel →