The llama.cpp project has added support for the embeddingGemma2 model, which handles text, vision, and audio inputs. This update is implemented via pull request #30054.