vLLM Adds Day-0 Support for Google's EmbeddingGemma 2 Multimodal Embeddings
Original title🎉 Excited to support EmbeddingGemma 2 from @GoogleDeepMind on day 0!
AISummary
vLLM announced day-0 support for EmbeddingGemma 2 from Google DeepMind, a bidirectional omni-modal embedding model that maps text, image, audio, video, and interleaved inputs into one vector space.
Users can try it with the latest vLLM nightly build using the command vllm serve google/embeddinggemma-2 --runner pooling. The quoted Google post says the model is built on the Gemma 4 architecture and released under Apache 2.0.
Source: vLLM · x.comPublished · added here