Google 发布 EmbeddingGemma 2:基于 Gemma 4 的原生多模态嵌入模型
Google 开源了 EmbeddingGemma 2,一个模型同时嵌入文本、图片、音频和视频,还支持维度裁剪,做 RAG 检索的可以直接试试。
Google 发布 EmbeddingGemma 2,是其首个原生多模态嵌入模型,基于 Gemma 4 构建,采用 Apache 2.0 许可证开源。模型支持 100+ 语言,可将文本、代码、图像、音频和视频嵌入到同一向量空间。共提供 4 个尺寸:740M 全模态、440M 文本+视觉、570M 文本+音频和 270M 纯文本,上下文长度为 8192。在 MTEB Code 基准上提升 14%,支持 Matryoshka 嵌入,可裁剪为 768d、512d、256d 或 128d。模型已在 Hugging Face 上架,可用 Sentence Transformers 和 LiteRT-LM 加载。
EmbeddingGemma 2 is out! Our first native multimodal embeddings built on Gemma 4 under Apache 2.0. 🤗 Embeds 100+ languages, code, images, audio, and video into one vector. 🧮 8,192 context 3️⃣ 740M omni, 440M text+vision, 570M text+audio, or 270M text-only 📈 +14% jump on MTEB Code 🪆 Matryoshka: 768d, 512d, 256d, or 128d 🤗 Available in Sentence Transformers and LiteRT-LM 🔓 Apache huggingface.co/google/embeddi… ueSDmb 💬 5 🔄 4 ❤️ 13 👀 632 📊 7 ⚡