1 paper
Yifu Qiu, Paul-Ambroise Duquenne, Holger Schwenk
We introduce V-SONAR, a vision-language embedding space extended from the text-only embedding space SONAR (Omnilingual Embeddings Team et al., 2026), which supports 1500 text langu…