On multimodal retrieval, the 9B scores 65.2% on ViDoRe v3, beating nemotron-colembed-v2-8b
antoine_chaffin · x · 2026-10-08
On multimodal retrieval, the 9B scores 65.2% average on ViDoRe v3 with images, outperforming nemotron-colembed-v2-8b and trailing only vision-specific EVIE at a much smaller embedding dimension. The 0.6B scores 62.3%, beating nemotron-colembed-v2-4b and qwen3-vl-embed-8b, and matching topk's 2B model.
Related event: Perplexity Open-Sources Multimodal Embedding Models pplx-embed-v2-late(38 posts)→
More from Models
- ChatGPT Work mode vs Codex: same quota, far more tasks done per 5-hour window — sasik520 · 2026-10-08
- Claude Haiku 5.5 ships with huge jumps: OSWorld 15.7%→72.4%, beats GPT-6 Luna across the board — mark_k · 2026-10-08
- OpenAI on GPT-6 Intelligent UI: the hard part is knowing when to show it — btibor91 · 2026-10-08
- Burkov slams watermarking in paid LLM outputs: 'I paid for this text' — burkov · 2026-10-08
- Epoch's InnovationEval: AI agents still far from producing real research innovations — Afinetheorem · 2026-10-08
- 113 decision models in 3 weeks: 70 built on Qwen, sub-cent per call — jonathanmalkin · 2026-10-08