跳到正文
Google Developers Blog·· 3 天前精选AI 评分76

Google DeepMind 发布端侧多模态嵌入模型 EmbeddingGemma 2

Bring multimodal semantic search to the edge with EmbeddingGemma 2

AI 导读

Google DeepMind 发布开放权重多模态嵌入模型 EmbeddingGemma 2,可将文本、图像、视频帧和音频映射到统一向量空间,参数量 740M,文本权重最低约 191MB 内存,完整多模态模型在 Google Pixel 11 Pro 上约 567MB。

推荐理由

EmbeddingGemma 2 把文本、图像、视频帧和音频映射到同一向量空间,并给出端侧内存与延迟数据,可据此判断本地多模态检索的落地边界。

来源:Google Developers Blog · developers.googleblog.com