10 citations · 10 across the 1 of their papers we have counts for
2 papers
cs.CV2021★ 10 cited
Step-Wise Hierarchical Alignment Network for Image-Text Matching
Zhong Ji, Kexin Chen, Haoran Wang
Image-text matching plays a central role in bridging the semantic gap between vision and language. The key point to achieve precise visual-semantic alignment lies in capturing the…
cs.CV2020
Consensus-Aware Visual-Semantic Embedding for Image-Text Matching
Haoran Wang, Ying Zhang, Zhong Ji +2
Image-text matching plays a central role in bridging vision and language. Most existing approaches only rely on the image-text instance pair to learn their representations, thereby…