Google releases EmbeddingGemma 2 — open 740M multimodal on-device embedding model
EmbeddingGemma 2 is a 740M-parameter, Apache 2.0–licensed multimodal embedding model from Google that natively maps text, images, audio, video and code into a single embedding space and is optimized for on-device use. It is built on the Gemma 4 architecture, supports modular encoder configurations, an 8K-token context window, Matryoshka Representation Learning for configurable vector sizes, and is available on Hugging Face and Kaggle.