3 citations · 4 across the 6 of their papers we have counts for
1 paper · 1 filter
Jinyin Wang, Haijing Zhang, Yihao Zhong +3
Image-text matching is a key multimodal task that aims to model the semantic association between images and text as a matching relationship. With the advent of the multimedia infor…