Google Releases EmbeddingGemma 2: 740M Open Multimodal Embedding Model Running on 0.5GB RAM

gnukeith · x · 2026-10-07

Google DeepMind released EmbeddingGemma 2, its first natively multimodal open embedding model for on-device use: 740M total parameters (270M text + 170M vision + 300M audio) under Apache 2.0, unifying code, images, audio, and video in a shared 768-dimensional space across 100+ languages, runnable locally on 0.5GB of RAM. Unsloth provides GGUF quants and a training guide; the model card includes 768d benchmark results with and without vector truncation.

Related event: Google releases open-source multimodal EmbeddingGemma 2(27 posts)→

Original post →

More from Models

Models channel →