EmbeddingGemma 2

Google DeepMind🇺🇸 United States
active

Version History

2major

Google released EmbeddingGemma 2, a 740M-parameter natively multimodal embedding model on the Gemma 4 architecture under Apache 2.0. Google says quantized versions need about 191MB (text-only) to 567MB (full multimodal) of active RAM on a Pixel 11 Pro.

Coverage

model release

Google releases EmbeddingGemma 2: 740M-parameter multimodal embedding model under Apache 2.0

Google announced EmbeddingGemma 2, a 740M-parameter natively multimodal embedding model built on the Gemma 4 architecture and released under Apache 2.0. Google says the quantized model needs about 191MB of active RAM for text-only weights and about 567MB for the full multimodal model on a Pixel 11 Pro. Google also launched a Mac app, AI Edge Foresight, to demonstrate it.

3 min read