0.6B visual embedding model rivals 8B models, keeps natural image search strong

antoine_chaffin · x · 2026-10-08

The author shares retrieval benchmark results for their visual embedding models. On ViDoRe (visual document retrieval), their 0.6B model is competitive with much larger models, close to Qwen3-VL-Embed-8B on MIRACL-Vision, while their 9B outperforms all compared models except Gemini-Embedding-2.\n\nOn PPLX-Q2I, an internal benchmark covering visual documents and natural images from production logs, both models beat Qwen3-VL-Embedding-8B by a wide margin, with the 9B close behind Gemini-Embedding-2. The author notes that although document retrieval is the primary use case, strong document retrieval performance does not come at the cost of natural image search.

Related event: Perplexity open-sources pplx-embed-v2-late retrieval models(36 posts)→

Original post →

More from Models

Models channel →