Perplexity Open-Sources Deep Research Benchmark WANDR

AccBalanced · x · 2026-07-15

Perplexity has open-sourced its internal benchmark WANDR, used to build its deep and wide research capabilities.

The quoted content mentions that this benchmark was utilized for the J-space reproduction around GLM 5.2, training reward models, reducing hallucinations via RL, and evaluating models on tasks like cancer prediction.

The core takeaway is that Perplexity has publicly released an internal evaluation benchmark designed to build "deep research" capabilities, demonstrating its application in training and experimental pipelines.

Related event: Perplexity open-sources internal research benchmark WANDR(9 posts)→

Original post →

More from Research

Research channel →