Tencent open-sources WeMM multimodal embedding models, 9B ranks first on dual benchmarks
aigclink · x · 2026-08-26
Tencent's WeChat Vision team open-sourced the WeMM-Embedding family of universal multimodal embedding models, available in 2B, 4B, and 9B sizes. The models support text, image, video, and interleaved inputs for unified retrieval. The 9B version ranked first on both MMEB-v2 and MMEB-v3 benchmarks and is now available on GitHub and Hugging Face.
Related event: Tencent Open-Sources WeMM-Embedding Multimodal Models, 9B Tops MMEB v2/v3(7 posts)→
More from Models
- Zai Releases Open-Weight GLM-5.3 Coding Model Amid Safety Criticism — dhadfieldmenell · 2026-08-28
- GLM-5.3 Flash High Reasoning Live on HF Providers; Devs Call It Opus 4.8-Class — _akhaliq · 2026-08-28
- Gemini 1.5 Flash inference speeds may exceed 300 tok/sec — Sentdex · 2026-08-28
- Google open-sources TimesFM for zero-shot forecasting of trends — mdancho84 · 2026-08-28
- Hands-on: Tencent Hy4 preview shows significant research gains over Hy3 — ShunyuYao12 · 2026-08-28
- Hy4 Preview builds complex Three.js temple in 87 minutes — ShunyuYao12 · 2026-08-28