2 citations · 2 across the 1 of their papers we have counts for
1 paper
Xiaohan Zou, Changqiao Wu, Lele Cheng +1
Most existing methods in vision-language retrieval match two modalities by either comparing their global feature vectors which misses sufficient information and lacks interpretabil…