• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
Technology
Announcement

Google DeepMind announces EmbeddingGemma 2 for on-device multimodal embeddings

Google says EmbeddingGemma 2 is its first natively multimodal open model for on-device embeddings, built on Gemma 4 to map text, code, images, audio, and video into a shared space locally.

Google DeepMindGD
GoogleGO
TechmemeTE
15 Sources, 3d ago, first seen 3d ago

TLDR

Google DeepMind announced EmbeddingGemma 2, which Google says is its first natively multimodal open model for on-device embeddings. Google says the Gemma 4-based model can map text, code, images, audio, and video into one shared space for local search and retrieval, but the specs and performance claims in the announcement come from Google’s own posts.

Combined views

1.8M

15 Sources, first seen 3d ago

7.5K likes360 comments3.8K saves963 reposts

Combined views

1.8M

15 Sources, first seen 3d ago

7.5K likes360 comments3.8K saves963 reposts

Google DeepMind announced EmbeddingGemma 2, which Google says is its first natively multimodal open model for on-device embeddings. Google says the model goes beyond text by putting code, images, audio, and video into a shared embedding space.

Featured Source

Google's pitch is that developers can use that shared representation to build local search and retrieval tools across different kinds of files. In practical terms, the company says EmbeddingGemma 2 is optimized to run locally on mobile and desktop, including offline uses such as finding a specific clip in audio or video with a text query.

What Google says the model is for

Google says EmbeddingGemma 2 is built on the Gemma 4 architecture and released under an Apache 2.0 license. It also says the model can work alongside Gemma 4 in an on-device retrieval-augmented generation setup, where EmbeddingGemma 2 retrieves local files and Gemma 4 reasons over them.

That makes this more about indexing and retrieval than about a standalone chatbot. Embedding models turn inputs into numerical representations that software can compare and search, which is why Google is framing this around local files, privacy, and offline use.

The specs in Google's announcement

In the announcement thread, Google says EmbeddingGemma 2 has 740 million parameters and uses from about 191 MB to 567 MB of active RAM. Google also says the model has an 8K context window, which it describes as four times larger than the first generation.

For multimodal inputs, Google says one pass can handle up to 5.5 minutes of audio, 29 images, or 58 video frames.

Those details help explain the intended role. Google is positioning EmbeddingGemma 2 as a relatively small open model for organizing and searching local media across formats, rather than as a large general-purpose model.

The main limitation in the supplied evidence is that the substantive details about capabilities, memory use, and intended applications come from Google's own announcement posts. That supports reporting the launch and Google's stated specs, but not independently confirming how the model performs in broader real-world use.

Still, the release shows where Google wants the Gemma line to expand: from text-focused embeddings to a single local model that can help apps search and retrieve several kinds of data without sending them to the cloud.

Sentiment

Positive50.1%49.9%Negative

Summary

Positive accounts welcomed EmbeddingGemma 2 for its strong on-device performance and private multimodal embeddings, while negative replies questioned Google's ad-driven data practices and noted competitors like Jina still lead.

Based on 102 sentiment-bearing replies from 73 accounts across 3 conversations.

Useful links

Google for Developers · YouTube

Introducing EmbeddingGemma 2: An open model for natively multimodal embeddings

Omar Kamal · YouTube

EmbeddingGemma 2 Explained: One Embedding Space for Text, Images, Video & Audio

Prompt Engineering · YouTube

Embedding Gemma 2: On-Device Multimodal RAG Made Easy

Sentiment

Positive50.1%49.9%Negative

Summary

Positive accounts welcomed EmbeddingGemma 2 for its strong on-device performance and private multimodal embeddings, while negative replies questioned Google's ad-driven data practices and noted competitors like Jina still lead.

Based on 102 sentiment-bearing replies from 73 accounts across 3 conversations.

Related Videos

  • Introducing EmbeddingGemma 2: An open model for natively multimodal embeddingsGoogle for Developers · YouTube
  • EmbeddingGemma 2 Explained: One Embedding Space for Text, Images, Video & AudioOmar Kamal · YouTube
  • Embedding Gemma 2: On-Device Multimodal RAG Made EasyPrompt Engineering · YouTube

Related Videos

  • Introducing EmbeddingGemma 2: An open model for natively multimodal embeddingsGoogle for Developers · YouTube
  • EmbeddingGemma 2 Explained: One Embedding Space for Text, Images, Video & AudioOmar Kamal · YouTube
  • Embedding Gemma 2: On-Device Multimodal RAG Made EasyPrompt Engineering · YouTube

Related

Google unveils universal Gemini agent for enterprise

Google says the agent can create sub-agents for multi-step tasks and maintain context across devices.

Google says it launched a test satellite with four TPUs for Project Suncatcher

Google says Project Suncatcher is a moonshot testing whether machine learning infrastructure could one day run in space, including whether its TPU hardware can keep working through radiation, thermal extremes, and other orbital hazards.

An infographic against a black background depicting the Earth below states that a rocket trip to low Earth orbit takes about 10 minutes, exposing spacecraft carrying TPUs to acceleration loads up to 10x gravity, and individual TPU chips to forces up to 50–100 g.
Google Drive and Docs reportedly now support Markdown files natively

Google says users can preview .md files in Drive, then open them in Docs to edit, comment on, and collaborate on them without converting the file into a Google Doc.

15 Sources

Google DeepMind@GoogleDeepMindMeet EmbeddingGemma 2, our first natively multimodal open model for on-device embeddings. It expands beyond text to unify code, images, audio, and video in a shared space. 🧵3d
Google@GoogleWe’re releasing EmbeddingGemma 2, our first natively multimodal open model engineered for on-device embeddings. Built on the Gemma 4 architecture and released under an Apache 2.0 license, it goes beyond text to unify images, video, audio, and code in a single embedding space.3d
Techmeme@TechmemeGoogle DeepMind launches EmbeddingGemma 2, a 740M-parameter model to map code, images, video, and audio in a shared embedding space, under an Apache 2.0 license (Google) (Visit Techmeme dot com for the link and full context!)3d
The New Stack@thenewstackGoogle's EmbeddingGemma 2 brings text, code, image, video and audio search to one on-device model, so developers can skip captioning and transcription. https://thenewstack.io/google-embeddinggemma-multimodal-search/?taid=6ac55348a01597000145b52f&utm_campaign=trueanthem&utm_medium=social&utm_source=twitter3d
RuntimeWire 🏴‍☠️@runtimewirehttps://runtimewire.com/article/google-embeddinggemma-2-on-device-multimodal-search3d
SiliconANGLE@SiliconANGLEGoogle expands EmbeddingGemma beyond text to images, audio and video https://ift.tt/mDRyPBZ3d
Emma Scharfman@EmmaScharfmannEmbeddingGemma made it possible for so many AI for Science projects to get started: 🩺 Embedding Gemma 300m Medical 🌍 Geo-Gemma 🎗️ Oncology-Gemma 📚 EmbeddingGemma PubMed EmbeddingGemma 2, which released and made available on Hugging Face yesterday, opens the door to so many new projects in science! 🚀 🧬 🌎 https://huggingface.co/google/embeddinggemma-22d
Hugging Face@huggingfaceRT @EmmaScharfmann: EmbeddingGemma made it possible for so many AI for Science projects to get started: 🩺 Embedding Gemma 300m Medical 🌍 Ge…2d
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    GoogleGoogle DeepMind

    15 Sources

    Google DeepMind@GoogleDeepMindMeet EmbeddingGemma 2, our first natively multimodal open model for on-device embeddings. It expands beyond text to unify code, images, audio, and video in a shared space. 🧵3d
    Google@GoogleWe’re releasing EmbeddingGemma 2, our first natively multimodal open model engineered for on-device embeddings. Built on the Gemma 4 architecture and released under an Apache 2.0 license, it goes beyond text to unify images, video, audio, and code in a single embedding space.3d
    Techmeme@TechmemeGoogle DeepMind launches EmbeddingGemma 2, a 740M-parameter model to map code, images, video, and audio in a shared embedding space, under an Apache 2.0 license (Google) (Visit Techmeme dot com for the link and full context!)3d
    The New Stack@thenewstackGoogle's EmbeddingGemma 2 brings text, code, image, video and audio search to one on-device model, so developers can skip captioning and transcription. https://thenewstack.io/google-embeddinggemma-multimodal-search/?taid=6ac55348a01597000145b52f&utm_campaign=trueanthem&utm_medium=social&utm_source=twitter3d
    RuntimeWire 🏴‍☠️@runtimewirehttps://runtimewire.com/article/google-embeddinggemma-2-on-device-multimodal-search3d
    SiliconANGLE@SiliconANGLEGoogle expands EmbeddingGemma beyond text to images, audio and video https://ift.tt/mDRyPBZ3d
    Emma Scharfman@EmmaScharfmannEmbeddingGemma made it possible for so many AI for Science projects to get started: 🩺 Embedding Gemma 300m Medical 🌍 Geo-Gemma 🎗️ Oncology-Gemma 📚 EmbeddingGemma PubMed EmbeddingGemma 2, which released and made available on Hugging Face yesterday, opens the door to so many new projects in science! 🚀 🧬 🌎 https://huggingface.co/google/embeddinggemma-22d
    Hugging Face@huggingfaceRT @EmmaScharfmann: EmbeddingGemma made it possible for so many AI for Science projects to get started: 🩺 Embedding Gemma 300m Medical 🌍 Ge…2d
    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet