Close Menu
NCIJ Network NCIJ Network
    What's Hot

    Tennessee’s Botched Death Row Execution Could Happen Anywhere — ProPublica

    October 6, 2026

    How to Avoid Disaster in the Next Iran War

    October 6, 2026

    Why Flavio Bolsonaro outperformed the polls in Brazil’s election | Corruption News

    October 6, 2026
    Facebook X (Twitter) Instagram
    Trending
    • Tennessee’s Botched Death Row Execution Could Happen Anywhere — ProPublica
    • How to Avoid Disaster in the Next Iran War
    • Why Flavio Bolsonaro outperformed the polls in Brazil’s election | Corruption News
    • Les renforts d’Europol en IA suscitent des inquiétudes sur la protection de la vie privée
    • Don’t mention the Iron Dome: comedy genius Kemi on warpath at Tory conference | John Crace
    • Kemi Badenoch is the most rightwing leader in Tory history. Isn’t it time she got some actual scrutiny? | Owen Jones
    • Best Amazon Prime Day Vacuum Deals: Dyson, Shark, and Robot Vacuums (2026)
    • Google DeepMind Releases EmbeddingGemma 2, a 740M Open Multimodal Embedding Model Built on Gemma 4
    • About
      • Our Team
      • Editorial Policy
      • Editorial Independence
      • International Support
    • Trust & Standards
      • AI Usage Policy
      • Conflict of Interest Policy
      • Corrections Policy
      • Ethics Policy
      • Fact-Checking Policy
      • Source Protection
    • Get Involved
      • Guide for Sources
      • Support Independent Journalism
    • Legal
      • Cookie Policy
      • Privacy Policy
      • Terms of Use
    Facebook X (Twitter) Instagram
    NCIJ Network NCIJ Network
    Tuesday, October 6
    • Home
    • World
    • Ai
    • Business
    • Politics
    • Health
    • Crypto
    • Science
    • Technology
    • Cybersecurity
    • Defense & Security
    • Economy
    • Energy
    • Europe
    • More
      • Fact Check
      • Investigations
      • Opinion & Analysis
      • Environment
    NCIJ Network NCIJ Network
    Home»Artificial Intelligence

    Google DeepMind Releases EmbeddingGemma 2, a 740M Open Multimodal Embedding Model Built on Gemma 4

    NCIJ NETWNCIJ NETWORKBy NCIJ NETWNCIJ NETWORKOctober 6, 2026 Artificial Intelligence No Comments4 Mins Read
    Share
    Facebook Twitter LinkedIn Pinterest Email

    Google DeepMind has released EmbeddingGemma 2, an open model that embeds text, code, images, video and audio into one 768-dimensional space. It has 740M parameters, an 8K token context window and an Apache 2.0 license. It targets on-device search, classification and privacy-first RAG. This article analyzes, compares and showcase how EmbeddingGemma 2 fits in the space.

    Deployable today? Yes. Weights are live on Hugging Face and Kaggle, with Ollama, llama.cpp GGUF and LiteRT builds available now.

    What an Embedding Model Does

    An embedding model converts content into a vector of numbers that captures meaning. Similar items land close together, so they are easy to search and compare. In a RAG pipeline, these vectors let an LLM retrieve fresh information it was not trained on. Generating embeddings locally keeps data on the device, cuts latency and works offline.

    One Vector Space for Every Modality

    EmbeddingGemma 2 is built on the Gemma 4 architecture. A text query can retrieve a photo. A voice memo can retrieve a video clip. Interleaved inputs, like a product listing with text, images and a demo video, produce a single embedding.

    The design is modular. It has three parts:

    • Text and code backbone: 270M parameters (130M transformer plus 140M embedder)
    • Vision encoder: 170M parameters, optional
    • Audio encoder: 300M parameters, optional

    Developers load only what they need: 270M for text, 440M for text and vision, 570M for text and audio, or 740M for everything. All setups share one vector space. A query embedded with the text-only setup can match documents embedded by the full model.

    The context window is 8,192 tokens, 4x larger than version 1. That fits about 29 images, 58 video frames or 5.5 minutes of audio.

    Benchmarks

    Google research team reports leading scores among sub-1B multimodal embedders on MTEB Code and MAEB. Full-precision results at 768 dimensions:

    Benchmark EmbeddingGemma 2 EmbeddingGemma 1
    MTEB multilingual v2 61.36 61.15
    MTEB Code v1 78.68 68.76
    MIEB lite (image) 64.64 n/a
    MMEB v2 overall 59.01 n/a
    MSEB retrieval (sound) 69.54 n/a
    MAEB (audio) 49.39 n/a

    Source: EmbeddingGemma 2 model card

    Code retrieval gains 9.92 points, roughly 14%. Multilingual text quality holds steady. Bigger models still lead some boards. Qwen3-VL-Embedding-2B reports 73.2 on its own MMEB-V2 run, with about 2.7x the parameters and no audio support.

    Built for Phones and Laptops

    With quantization on a Pixel 11 Pro, active RAM is about 191MB for text-only weights. The full multimodal model needs about 567MB. Quantization-aware training compresses weights to INT4 and INT8. The Google AI Edge team measured 37.3 ms per image on a MacBook M5 Pro GPU, using a 70-token vision budget.

    Matryoshka Representation Learning (MRL) lets developers truncate vectors to 512, 256 or 128 dimensions. Moving from 768 to 128 dimensions cuts storage up to 6x. At 256 dimensions, MTEB multilingual only slips from 61.36 to 60.41. At 128 dimensions, MMEB drops to 45.65, so Google recommends 128d mainly for text-only workloads.

    Interactive Explainer

    740M Built DeepMind Embedding EmbeddingGemma Gemma Google model Multimodal open Releases
    NCIJ NETWNCIJ NETWORK
    • Website

    Keep Reading

    Use Gemini for free? You’ll soon be limited to its weakest AI model

    Mistral’s new Le Chonk model brings AI cybersecurity to your business – and you control it

    Google Pauses OSS Product Bug Bounty Rewards After Surge in Invalid Automated Reports

    Reka Releases Rho-1: A 19B Omni-Reasoning Model That Understands, Generates Video and Outputs Robot Actions in One

    Beyond Domain-Specific World Models: JEPA-Anything Uses 1 Recipe for 7 Fields

    Reflection AI Introduces Beam: A 501B Open-Weight MoE Model With 23B Active Parameters for Coding and Agentic Workloads

    Add A Comment
    Leave A Reply Cancel Reply

    Editors Picks

    Tennessee’s Botched Death Row Execution Could Happen Anywhere — ProPublica

    October 6, 2026

    How to Avoid Disaster in the Next Iran War

    October 6, 2026

    Why Flavio Bolsonaro outperformed the polls in Brazil’s election | Corruption News

    October 6, 2026

    Les renforts d’Europol en IA suscitent des inquiétudes sur la protection de la vie privée

    October 6, 2026
    Latest Posts

    4 Best Compression Boots: Therabody, Hyperice, and More (2026)

    August 9, 2026

    Former Iraqi provincial governor arrested as graft crackdown continues | Corruption News

    August 9, 2026

    The culture surrounding ‘ideal’ childbirth has to evolve | Childbirth

    August 9, 2026

    Subscribe to News

    Get the latest sports news from NewsSite about world, sports and politics.

    NCIJ Network is an independent digital news platform delivering trusted investigative journalism, European and global news, in-depth analysis, and fact-based reporting with accuracy, transparency, and integrity.

    Facebook X (Twitter) Instagram Pinterest YouTube

    Tennessee’s Botched Death Row Execution Could Happen Anywhere — ProPublica

    October 6, 2026

    How to Avoid Disaster in the Next Iran War

    October 6, 2026

    Why Flavio Bolsonaro outperformed the polls in Brazil’s election | Corruption News

    October 6, 2026

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    Type above and press Enter to search. Press Esc to cancel.