1 citations · 1 across the 1 of their papers we have counts for
1 paper · 1 filter
Hanqi Yan, Xiangxiang Cui, Lu Yin +4
The success of vision-language models is primarily attributed to effective alignment across modalities such as vision and language. However, modality gaps persist in existing align…