Abstract
EmbeddingGemma, a lightweight text embedding model based on Gemma 3, achieves state-of-the-art performance with fewer parameters through encoder-decoder initialization, geometric embedding distillation, and spread-out regularization.
We introduce EmbeddingGemma, a new lightweight, open text embedding model based on the Gemma 3 language model family. Our innovative training recipe strategically captures knowledge from larger models via encoder-decoder initialization and geometric embedding distillation. We improve model robustness and expressiveness with a spread-out regularizer, and ensure generalizability by merging checkpoints from varied, optimized mixtures. Evaluated on the Massive Text Embedding Benchmark (MTEB) across multilingual, English, and code domains, EmbeddingGemma (300M) achieves state-of-the-art results. Notably, it outperforms prior top models, both proprietary and open, with fewer than 500M parameters, and provides performance comparable to models double its size, offering an exceptional performance-to-cost ratio. Remarkably, this lead persists when quantizing model weights or truncating embedding outputs. This makes EmbeddingGemma particularly well-suited for low-latency and high-throughput use cases such as on-device applications. We provide ablation studies exploring our key design choices. We release EmbeddingGemma to the community to promote further research.
Get this paper in your agent:
hf papers read 2509.20354
curl -LsSf https://hf.co/cli/install.sh | bash
Models citing this paper 291
![]()
google/embeddinggemma-300m
Sentence Similarity •
0.3B •
Updated
•
2.5M
•
1.86k
![]()
google/embeddinggemma-300m-qat-q4_0-unquantized
Sentence Similarity •
0.3B •
Updated
•
208
•
47
![]()
google/embeddinggemma-300m-qat-q8_0-unquantized
Sentence Similarity •
0.3B •
Updated
•
3.07k
•
44
![]()
lmstudio-community/embeddinggemma-300m-qat-GGUF
0.3B •
Updated
•
1.86k
•
10
Datasets citing this paper 1
Vano04/laions-got-talent-enhanced-precomputed-en
Viewer •
Updated
•
1.12M
•
136