Ollama hosts Google's EmbeddingGemma 2, a 740M multimodal embedding model for on-device use

ollama · x · 2026-10-07

Ollama now serves EmbeddingGemma 2, Google DeepMind's open multimodal embedding model: 740M total params (270M text + 170M vision + 300M audio encoders) mapping text, images, video, and audio into a unified 768-dim vector space with 256K context. Sizes range from 270m (378MB) to 740m (1.3GB), designed for low-latency on-device semantic applications.

Original post →

More from Models

Models channel →