跳到正文
原文
Google Developers · AI 筛选·· 2 小时前精选AI 评分76

Google 发布端侧多模态嵌入模型 EmbeddingGemma 2

Bring multimodal semantic search to the edge with EmbeddingGemma 2

AI 导读

Google DeepMind 正式推出开源多模态嵌入模型 EmbeddingGemma 2,可将文本、图像、视频帧和音频原生映射至统一向量空间。该模型参数量为 740M,纯文本权重仅需约 191MB 活跃内存,全多模态在 Pixel 设备上仅占用约 567MB,支持零样本意图路由与端侧毫秒级决策。

推荐理由

该模型将文本、视觉与音频统一映射至单一向量空间,为端侧离线语义检索与零样本决策提供了轻量化实现方案。

来源:Google Developers · AI 筛选 · developers.googleblog.com