Average Ratings 0 Ratings
Average Ratings 0 Ratings
Description
EmbeddingGemma 2 is a versatile and lightweight multimodal embedding model that facilitates the mapping of text, code, images, video, and audio into a unified embedding space, which is useful for applications such as search, retrieval, classification, routing, and RAG. It is constructed on the Gemma 4 framework and distributed under the Apache 2.0 license, featuring an impressive 740 million parameters while being fine-tuned for efficient on-device inference. Its flexible architecture allows for the use of only 270 million parameters for text-centric tasks, and it includes additional vision and audio encoders for comprehensive multimodal capabilities. Furthermore, the innovative Matryoshka Representation Learning technique enables developers to compress output vectors from 768 dimensions down to 512, 256, or even 128 dimensions, effectively reducing the storage and memory demands for local vector databases. The model is equipped with an 8K-token context window, providing the capability to handle up to 5.5 minutes of audio, 29 images, 58 video frames, or various combinations of these inputs seamlessly on local hardware. This adaptability makes it particularly valuable for developers seeking to enhance their applications with rich multimedia integration.
Description
Introducing Gemma 3n, our cutting-edge open multimodal model designed specifically for optimal on-device performance and efficiency. With a focus on responsive and low-footprint local inference, Gemma 3n paves the way for a new generation of intelligent applications that can be utilized on the move. It has the capability to analyze and respond to a blend of images and text, with plans to incorporate video and audio functionalities in the near future. Developers can create smart, interactive features that prioritize user privacy and function seamlessly without an internet connection. The model boasts a mobile-first architecture, significantly minimizing memory usage. Co-developed by Google's mobile hardware teams alongside industry experts, it maintains a 4B active memory footprint while also offering the flexibility to create submodels for optimizing quality and latency. Notably, Gemma 3n represents our inaugural open model built on this revolutionary shared architecture, enabling developers to start experimenting with this advanced technology today in its early preview. As technology evolves, we anticipate even more innovative applications to emerge from this robust framework.
API Access
Has API
Yes
API Access
Has API
No
Integrations
Gemini
No
Gemini Enterprise
No
Gemini Enterprise Agent Platform
No
Gemini Nano
No
Gemma
No
Google AI Edge
No
Google AI Edge Gallery
No
Google AI Studio
No
Google Cloud Platform
No
Hugging Face
No
Integrations
Gemini
Yes
Gemini Enterprise
Yes
Gemini Enterprise Agent Platform
Yes
Gemini Nano
Yes
Gemma
Yes
Google AI Edge
Yes
Google AI Edge Gallery
Yes
Google AI Studio
Yes
Google Cloud Platform
Yes
Hugging Face
Yes
Pricing Details
No price information available.
Free Trial
No
Free Version
No
Pricing Details
No price information available.
Free Trial
Yes
Free Version
No
Deployment
Web-Based
Yes
On-Premises
No
iPhone App
No
iPad App
No
Android App
No
Windows
No
Mac
No
Linux
No
Chromebook
No
Deployment
Web-Based
No
On-Premises
Yes
iPhone App
No
iPad App
No
Android App
No
Windows
Yes
Mac
Yes
Linux
Yes
Chromebook
No
Customer Support
Business Hours
No
Live Rep (24/7)
No
Online Support
Yes
Customer Support
Business Hours
Yes
Live Rep (24/7)
No
Online Support
Yes
Types of Training
Training Docs
Yes
Webinars
No
Live Training (Online)
No
In Person
No
Types of Training
Training Docs
Yes
Webinars
Yes
Live Training (Online)
No
In Person
Yes
Vendor Details
Company Name
Founded
1998
Country
United States
Website
blog.google/innovation-and-ai/technology/developers-tools/embeddinggemma-2/
Vendor Details
Company Name
Google DeepMind
Founded
2010
Country
United Kingdom
Website
deepmind.google/models/gemma/gemma-3n/
Product Features
Product Features
Alternatives
No Alternatives