
EmbeddingGemma 2 (2026): One Small Model That Puts Text, Images, Audio and Video in the Same Vector Space
Google DeepMind released EmbeddingGemma 2 on October 6, 2026. Using the official blog and the model card, this guide covers how it maps text, code, images, video and audio into one 768-dimensional space, how to pick between 270M and 740M effective sizes, Google's published benchmarks including the MTEB Code jump from 68.76 to 78.68, Matryoshka truncation down to 128 dimensions, how to run it with sentence-transformers, Ollama or llama.cpp, and the pitfalls that bite in practice such as never using float16.






























