Google Developers · AI 筛选·· 2 小时前精选AI 评分76
Google 发布端侧多模态嵌入模型 EmbeddingGemma 2
Bring multimodal semantic search to the edge with EmbeddingGemma 2
AI 导读
Google DeepMind 正式推出开源多模态嵌入模型 EmbeddingGemma 2,可将文本、图像、视频帧和音频原生映射至统一向量空间。该模型参数量为 740M,纯文本权重仅需约 191MB 活跃内存,全多模态在 Pixel 设备上仅占用约 567MB,支持零样本意图路由与端侧毫秒级决策。
推荐理由
该模型将文本、视觉与音频统一映射至单一向量空间,为端侧离线语义检索与零样本决策提供了轻量化实现方案。
来源:Google Developers · AI 筛选 · developers.googleblog.com