← r/LocalLLaMA
▲
171
+125
32👁
r/LocalLLaMA · u/Recoil42 · 3d ago

Introducing EmbeddingGemma 2: A best-in-class open model for natively multimodal embeddings | Google

https://huggingface.co/google/embeddinggemma-2

EmbeddingGemma 2 is an open multimodal embedding model built by Google DeepMind which maps text (incl. code), images, video, and audio inputs—and combinations thereof—into a single, unified 768-dimensional vector space. The model has 740M total parameters, combining a 270M parameter text model with modular vision (170M) and audio (300M) encoders.

Designed to run on consumer hardware such as mobile devices and laptops, EmbeddingGemma 2 delivers low-latency semantic representations for on-device applications, like search, retrieval-augmented generation (RAG), classification, and clustering.

171 0 171 10/6 17:47 10/8 23:02 UTC
scorecomments32 sightings
first seen 2026-10-06 17:47 UTClast seen 2026-10-08 23:02 UTCscore then 46score now 171gained +125sightings 32
open on reddit ↗ 💬 34 (+26)