1 citations · 1 across the 4 of their papers we have counts for
4 papers
Memory-Augmented Query Intent Understanding for Efficient Chat-based Image Retrieval
Xianke Chen, Daizong Liu, Yushuo Lou +5
Different from traditional text-to-image retrieval tasks, chat-based image retrieval allows the human-interactive system to iteratively clarify and refine user intent through multi…
Dynamic Adapter with Semantics Disentangling for Cross-lingual Cross-modal Retrieval
Rui Cai, Zhiyu Dong, Jianfeng Dong +1
Existing cross-modal retrieval methods typically rely on large-scale vision-language pair data. This makes it challenging to efficiently develop a cross-modal retrieval model for u…
Representation Alignment Contrastive Regularization for Multi-Object Tracking
Zhonglin Liu, Shujie Chen, Jianfeng Dong +2
Achieving high-performance in multi-object tracking algorithms heavily relies on modeling spatio-temporal relationships during the data association stage. Mainstream approaches enc…
Dual-view Curricular Optimal Transport for Cross-lingual Cross-modal Retrieval
Yabing Wang, Shuhui Wang, Hao Luo +5
Current research on cross-modal retrieval is mostly English-oriented, as the availability of a large number of English-oriented human-labeled vision-language corpora. In order to b…