Google 开源端侧多模态嵌入模型 EmbeddingGemma 2,并推出离线会议笔记应用
先了解这件事
Google DeepMind 开源了多模态嵌入模型 EmbeddingGemma 2,参数约 740M,基于 Gemma 4 架构,采用 Apache 2.0 协议,可在本地将文本、代码、图像、音频和视频映射到统一 embedding 空间,权重已在 Hugging Face 和 Kaggle 开源。Google 称该模型在多模态嵌入 benchmark 上优于体量两倍于它的竞品,并称是同类模型中最小的一款;相比前代,上下文窗口扩大到 8K token,MTEB (Code) 得分由 68.76 升至 78.68。 EmbeddingGemma 2 量化后在 Pixel 11 Pro 上仅文本权重约占 191MB 活动内存、完整多模态模型约占 567MB,可本地处理最多约 5.5 分钟音频、29 张图像、58 帧视频或上述模态的交错组合,无需 API key。 随后 Google 推出 Mac 端本地优先 AI 会议笔记应用 Google AI Edge Foresight,针对 Apple Silicon 优化,可完全离线运行,调用端侧 EmbeddingGemma 2 转录会议内容,并由 Gemma 4 驱动的助手进行问答。
AI 根据报道生成 · 20 小时前更新
事件进展
2 个进展
- 10月8日 21:28 · 2 篇报道Google 发布本地优先的 AI 笔记应用 AI Edge ForesightThe Verge · AI:Google 发布实验性离线 AI 笔记应用 AI Edge Foresight
- 10月7日 03:47 · 2 篇报道EmbeddingGemma 2 开源轻量多模态嵌入模型发布Google DeepMind:Google DeepMind 发布 EmbeddingGemma 2:面向端侧的多模态推理 embedding 模型
报道时间线
沿着报道,了解事件的不同侧面。
- The Verge · AIGoogle 发布实验性离线 AI 笔记应用 AI Edge Foresight
Google 发布实验性 AI 笔记应用 Google AI Edge Foresight,可在 macOS 上完全离线转录会议和录音,底层使用端侧 EmbeddingGemma 2 模型。
- TechCrunch · AIGoogle 发布本地优先的 AI 会议笔记应用 Google AI Edge Foresight
Google 推出面向 Mac 的本地优先 AI 会议笔记应用 Google AI Edge Foresight,可完全离线运行,调用端侧 EmbeddingGemma 2(740M 参数)转录会议内容,并借助 Gemma 4 驱动的助手进行问答。
- Google DeepMindGoogle DeepMind 发布 EmbeddingGemma 2:面向端侧的多模态推理 embedding 模型
Google DeepMind 发布 EmbeddingGemma 2,基于 Gemma 4 架构,采用 Apache 2.0 协议开源,参数规模约 740M,可将文本、代码、图像、音频和视频映射到统一 embedding 空间。
- The DecoderGoogle 开源 EmbeddingGemma 2,主打 740M 参数多模态嵌入
Google 发布开源多模态嵌入模型 EmbeddingGemma 2,拥有 740M 参数,支持文本、图像、视频、音频和代码转向量。Google 称它在多模态嵌入 benchmark 上表现优于体量两倍于它的竞品,并称这是同类模型中最小的一款;在 MTEB (Code) 上得分 78.68,较前代 68.76 提升近 10 分。
本事件热度走势
还没有足够的连续观测数据,暂不绘制趋势。