2 papers
cs.CV2025
Ambiguity-Aware and High-Order Relation Learning for Multi-Grained Image-Text Matching
Junyu Chen, Yihua Gao, Mingyuan Ge +1
Image-text matching is crucial for bridging the semantic gap between computer vision and natural language processing. However, existing methods still face challenges in handling hi…
cs.MM2025
Visual Semantic Description Generation with MLLMs for Image-Text Matching
Junyu Chen, Yihua Gao, Mingyong Li
Image-text matching (ITM) aims to address the fundamental challenge of aligning visual and textual modalities, which inherently differ in their representations, continuous, high-di…