Perplexity Releases Q2D-Web: 190M-Doc Benchmark for Agentic RAG Retrieval

_reachsumit · x · 2026-09-09

Perplexity AI introduced Q2D-Web, a large-scale benchmark for first-stage retrieval in agentic RAG, pairing a 190M-document web corpus with 70k agent-reformulated queries drawn from production user queries across ten languages. It addresses gaps in existing benchmarks, which either have huge corpora but few queries, or many queries but small corpora, and typically test human-written queries rather than machine reformulations. Three relevance judgment sets are provided (agent citations, production rankings, and a union augmented with LLM judgments). Benchmarks across 13 lexical, dense, and late-interaction retrievers show rankings are largely insensitive to judgment-set choice.

Original post →

More from Research

Research channel →