Tencent open-sources WeMM multimodal embedding models, 9B ranks first on dual benchmarks

aigclink · x · 2026-08-26

Tencent's WeChat Vision team open-sourced the WeMM-Embedding family of universal multimodal embedding models, available in 2B, 4B, and 9B sizes. The models support text, image, video, and interleaved inputs for unified retrieval. The 9B version ranked first on both MMEB-v2 and MMEB-v3 benchmarks and is now available on GitHub and Hugging Face.

Related event: Tencent Open-Sources WeMM-Embedding Multimodal Models, 9B Tops MMEB v2/v3(7 posts)→

Original post →

More from Models

Models channel →