Urgent.News

What's breaking now, across thousands of outlets.

AI

EmbeddingGemma 2

Yesterday, we unveiled EmbeddingGemma 2, a versatile model that extends text embeddings to code, images, video, and audio. Developed on the Gemma 4 architecture and released under the Apache 2.0 license, EmbeddingGemma 2 boasts 740 million parameters, making it ideal for on-device use. This model can locate specific video clips from voice memos, search audio recordings using text queries, and more, all processed by a single, multimodal model.

EmbeddingGemma 2 maintains the high-quality text performance of its predecessor while significantly improving code performance, achieving a 9.92-point increase in MTEB Code (from 68.76 to 78.68). This makes it an excellent choice for indexing local codebases, performing semantic code searches, and conducting coding agent retrievals.

In terms of performance, EmbeddingGemma 2 excels across various data types, setting a new standard for quality-per-parameter in sub-1B models and even surpassing some specialist models more than twice its size. For detailed evaluation metrics and model information, refer to the EmbeddingGemma 2 model card.

By generating embeddings locally, EmbeddingGemma 2 enhances data privacy, reduces pipeline latency, and enables developers to create cross-modal search and retrieval systems that operate entirely offline. When integrated with generative models such as Gemma 4, EmbeddingGemma 2 supports on-device Retrieval Augmented Generation (RAG) pipelines capable of understanding complex multimodal data.

The model's compatibility with the Gemma 4 architecture allows it to share the same text tokenizer and audio encoder, reducing the combined memory footprint when used alongside other models. To explore how to build on-device search and RAG systems with LiteRT, consult the Google AI Edge blog post.

EmbeddingGemma 2 has been developed in partnership with several developers to ensure seamless integration into your existing workflows. For a comprehensive guide on building on-device search and RAG systems using LiteRT, check out the developer guide, documentation, and other resources provided.

Written by urgent.news from Hacker News's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.

Also reported by 1 other outlet

Read the original at blog.google →

More in AI

CodeSmith How Ledgers and Snapshots Cure AI Debt

Ledger and Snapshots: The Debt AI Owes and the Regret Medicine Left for You Source version of CodeSmith : v0.5.0 (commit 3a74c82f ).

  • CodeSmith introduces Slop Ledger to track AI's technical residue
  • Ledger classifies slop into ten categories like compatibility shims and dead code
  • Four tools (append, query, update, export) manage the Slop Ledger

More from Tuesday 6 October →