Tencent open-sources EVIE visual document retrieval models, hitting 66.75 nDCG@10 on ViDoRe V3
jacek2023 · reddit · 2026-09-07
Tencent released EVIE-8B and EVIE-4.5B high-capacity visual document retrieval models on Hugging Face, scoring 66.75 and 66.02 nDCG@10 on ViDoRe V3. Key points:
- 8B model outputs 4096D per-token multi-vector embeddings preserving layout, typography, charts and table structures
- 4.5B uses Prefix-MRL: a single 2048D projection truncatable at runtime to 64–2048 dimensions without separate models
- Training-free HAC clustering compresses 750 tokens/page to 32 vectors, cutting index storage to 3.81 GiB per million pages
- EVIE-ARD distillation recipe reproduces student training from the 8B teacher; validated across 138 multilingual tasks on ViDoRe V1/V2/V3 and JinaVDR
More from Models
- 'Not AGI': User Spent $200 in 8 Hours on Astra, Found 3 of 4 Tasks Broken — Firm-Club-8334 · 2026-09-07
- Similarweb: ChatGPT's AI traffic share falls from 73.3% to 55.5% in 12 months — gaganghotra_ · 2026-09-07
- GPT-6 Astra (and Pro?) spotted on Simple-Bench leaderboard — From_Internets · 2026-09-07
- Testing Gemini as music understanders: Pro 3.1 solid, Flash models hallucinate sounds — teropa · 2026-09-07
- GPT-6 Astra posts 65.6% on ClockBench — Well_being1 · 2026-09-07
- Tencent Hunyuan ships Hy4 preview upgrade: same quality, fewer turns, lower token usage — xeophon · 2026-09-07