## Google 开源 EmbeddingGemma 2 原生多模态嵌入模型
Google DeepMind 10 月 6 日正式发布 EmbedGemma 第二代 —— EmbeddingGemma 2,新模型基于 Gemma 4 架构,参数规模约 3.08 亿至 7.4 亿,原生支持文本、代码、图像、视频与音频统一映射到同一 768 维向量空间,可直接在手机、笔记本、平板等端侧设备运行。
作为 Google 首款原生多模态嵌入模型,EmbeddingGemma 2 将传统需要分别处理的多模态检索合并到一次推理:开发者可在统一向量空间中跨文本、图像、视频与音频执行检索、分类与聚类任务,无需拼接多个单模态模型。新模型延续上一代「小而强」路线,参数量比同档竞品低一个数量级,但 MTEB 多语言榜单成绩接近闭源旗舰。
模型权重已在 Hugging Face 与 Kaggle Models 上线,LiteRT 与 Transformers、sentence-transformers 三大推理框架同步支持;并采用 Apache 2.0 商业友好协议发布,允许免费商用与再分发。Google 同步放出的开发者指南显示,EmbeddingGemma 2 在量化到 int4 后仍能保留 90% 以上检索精度,单条 query 延迟可压到 15ms 以内。
EmbeddingGemma 2 的开源恰逢 Gemini Embedding 2 商用版发布的同一天,Google 首次在「端侧小模型 + 云端旗舰」两侧同时铺开原生多模态嵌入能力,与 OpenAI text-embedding-3、Cohere Embed v3、Qwen3-Embedding 等形成正面竞争。
---
**参考来源:**
1. Google DeepMind 官方博客 — EmbeddingGemma 2 介绍页:https://deepmind.google/blog/embeddinggemma-2-an-open-lightweight-multimodal-embedding-model/
2. Google 开发者博客 — EmbeddingGemma 2: The Developer Guide:https://developers.googleblog.com/embeddinggemma-2-the-developer-guide/
3. 钜亨网 — Google 發表開源原生多模態嵌入模型 EmbeddingGemma 2:https://news.cnyes.com/news/id/6623578
4. 搜狐科技 — 谷歌发布 EmbeddingGemma 2:7.4 亿参数多模态嵌入模型:https://www.sohu.com/a/1084588616_130887
5. 雅虎新闻 — Google 推出首款原生多模态嵌入模型 Gemini Embedding 2:https://tw.news.yahoo.com/breaking-down-data-type-boundaries-google-launches-gemini-embedding-2-the-first-native-multimodal-embedding-model-163454511.html
6. Google AI for Developers — EmbeddingGemma 模型概览:https://ai.google.dev/gemma/docs/embeddinggemma?hl=zh-cn
7. Google AI for Developers — EmbeddingGemma 模型卡:https://ai.google.dev/gemma/docs/embeddinggemma/model_card?hl=zh-cn
8. DeepMind — Gemini Embedding 2 模型页:https://deepmind.google/models/gemini/embedding/







