跳到正文
RCreddit.com·
暂不在当前实时榜单

Introducing EmbeddingGemma 2: A best-in-class open model for natively multimodal embeddings | Google

AI 摘要

Google DeepMind has introduced EmbeddingGemma 2, an open multimodal embedding model. This model maps text, images, video, and audio inputs into a unified 768-dimensional vector space. With 740M parameters, it combines a 270M parameter text model with modular vision (170M) and audio (300M) encoders. Designed for consumer hardware, EmbeddingGemma 2 provides low-latency semantic representations for on-device applications such as search, RAG, classification, and clustering.

为什么是这条

This model is the first to natively support multimodal embeddings across text, images, video, and audio within a single 768-dimensional vector space.

时间与来源

时间显示为 UTC

显示时区:UTC

本地时区尚不可用,暂时显示 UTC。

收录当时偏移:UTC+02026年10月6日 23:00 UTC

收录
2026年10月6日 23:00
来源类型
开发者社区

本站未收录正文。

前往源站阅读 →
来源·reddit.com