← October 7, 2026 briefing

P1 Model Open weights GoogleHugging Face

Google releases EmbeddingGemma 2, an open embedding model that maps text, images, audio and video into one vector space

Announced October 6, 2026

What happened

Google DeepMind has released EmbeddingGemma 2, a lightweight open multimodal embedding model. According to the Hugging Face Transformers v5.19.0 release notes, it is based on the Gemma 4 architecture and encodes text, images, audio and video, alone or combined, into a shared 768-dimensional vector space. It uses Matryoshka Representation Learning, so embeddings can be truncated to 512, 256 or 128 dimensions. Unused vision and audio towers can be disabled at load time to save memory.

Why it matters

Transformers and Ollama (MLX) supported it on release day, so it can be used right away for cross-modal search and RAG. Simon Willison noted that its Apache 2.0 license reduces the risk of having to recompute embeddings.

Sources

More about Google