Ollama hosts Google's EmbeddingGemma 2, a 740M multimodal embedding model for on-device use
ollama · x · 2026-10-07
Ollama now serves EmbeddingGemma 2, Google DeepMind's open multimodal embedding model: 740M total params (270M text + 170M vision + 300M audio encoders) mapping text, images, video, and audio into a unified 768-dim vector space with 256K context. Sizes range from 270m (378MB) to 740m (1.3GB), designed for low-latency on-device semantic applications.
More from Models
- Testing Mistral Large 4 on FPS: 151K Reasoning Tokens Later, Still Rough — qtnx_ · 2026-10-07
- Many Mistral Large 4 failures traced to reasoning mode not being enabled — qtnx_ · 2026-10-07
- Early Opus 5.5 user says hype is overblown: shortcuts, wrong assumptions, sloppy work — haider1 · 2026-10-07
- TypeSafe's Jev model bets on machine-native intelligence over text-optimized LLMs — TWIML AI Podcast · 2026-10-07
- llm-mistral 0.16 adds reasoning model support for Mistral Large 4 — Simon Willison · 2026-10-07
- Reddit users grow frustrated with ChatGPT's over-refusals on innocuous prompts — Crixusgannicus · 2026-10-07