Matryoshka training lets embeddings truncate to 256d with a third of the storage

tomaarsen · x · 2026-10-07

Google's new multimodal embedding model is Matryoshka-trained, supporting 768/512/256/128 dimensions. At 256d, vectors take a third of the storage while MTEB multilingual v2 only drops from 61.36 to 60.41. Usage in Sentence-Transformers: model.encode(..., truncatedim=256, normalizeembeddings=True), re-normalizing after truncation. Official 768d full-precision results: 67.84 NDCG@5 on MMEB v2 visual documents, 50.67 Hit@1 on video, and 69.54 MRR@10 on MSEB audio retrieval.

Original post →

More from Multimodal

Multimodal channel →