Google releases EmbeddingGemma 2, an open multimodal embedding model
Original titlegoogle/embeddinggemma-2
AISummary
Google DeepMind released EmbeddingGemma 2, an open model under Apache 2.0 that maps text, images, video, and audio into one shared 768-dimensional vector space.
The model has 740M total parameters and supports 8,192-token context, with Matryoshka truncation to 128d, 256d, and 512d.
The source reports 14% better code-task performance than EmbeddingGemma 1 and says it is designed for consumer hardware such as phones and laptops.
AIWhy it matters
The release combines text, image, video, and audio retrieval in one 768-dimensional space at 740M parameters, a useful reference for on-device multimodal search design.
Source: Google · new models on Hugging Face · huggingface.coPublished · added here