Perplexity benchmarks 13 retrieval models: pplx-embed-v1-4b leads two of three categories
perplexity_ai · x · 2026-09-10
Perplexity released a technical report evaluating 13 retrieval models across three relevance sets, using Recall@1000 as the primary metric.
- pplx-embed-v1-4b leads Web Ranking (65.73) and Combined (69.11)
- Nemotron-3-Embed-8B leads Citation (61.68)
Anyone can request an evaluation by submitting a publicly available Hugging Face retrieval model via their form; full methodology is in the report.
More from Models
- Anthropic says Claude models accessed real systems during cyber evals; METR to investigate — scottleibrand · 2026-09-10
- Mathematician presents evidence OpenAI may have trained Astra on his Gromov soficity proof conversations — ValerioCapraro · 2026-09-10
- Letting GPT-6 Astra drive software UIs directly beats built-in agents, dev reports — brandon_galang · 2026-09-10
- Mathematician alleges OpenAI trained Astra on his unpublished work on Gromov's soficity conjecture — ValerioCapraro · 2026-09-10
- Users joke Claude is pushing them toward femboy personas — shakoistsLog · 2026-09-10
- GPT-6 'Astra' Does 34 Math Steps in Latent Space, 4x More Than Sol — MaartenBaert · 2026-09-10