Google launches EmbeddingGemma 2, an open 740M multimodal embedding model for on-device use
Recoil42 · reddit · 2026-10-07
Google DeepMind released EmbeddingGemma 2, an open multimodal embedding model that maps text (including code), images, video, and audio—individually or combined—into a single 768-dimensional vector space. The 740M-parameter model combines a 270M text model with modular vision (170M) and audio (300M) encoders. It's designed to run on consumer hardware like phones and laptops, targeting low-latency on-device search, RAG, classification, and clustering.
More from Models
- Paradigm releases tech report for Violetto, a 1B model trained from scratch for math — tensorqt · 2026-10-07
- Mistral's Le Chonk tops blind human code review among open models, second only to Opus 5 — qtnx_ · 2026-10-07
- DeepSeek V4.1 Flash hits 72.9% on ARC-AGI-2 at $0.13/task, costing 250% more — teortaxesTex · 2026-10-07
- Mistral claims Large 4 is one of the world's strongest AI models for cybersecurity — scaling01 · 2026-10-07
- Mistral Large 4 solves 18 of 19 CTF challenges in official speedrun with tool calls — MistralAI · 2026-10-07
- Community poll tiers AI labs: Anthropic and OpenAI frontier, Mistral and Amazon judged 3 generations behind — NathanpmYoung · 2026-10-07