Google 发布 EmbeddingGemma 2:740M 多模态嵌入模型,文本图像音视频统一向量空间

jacek2023 · reddit · 2026-10-06

Google DeepMind 发布开源多模态嵌入模型 EmbeddingGemma 2,已上架 Hugging Face(unsloth 提供 GGUF 版)。模型共 740M 参数,把文本(含代码)、图像、视频、音频统一映射到 768 维单一向量空间,专为手机、笔记本等消费级设备设计。

要点:

适用场景包括端侧搜索、RAG、分类与聚类。

所属事件:谷歌开源 EmbeddingGemma 2 多模态端侧嵌入模型(37 条相关)→

原文链接 →

「Infra」频道最新

更多「Infra」频道 AI 资讯 →