5 citations · 6 across the 3 of their papers we have counts for
3 papers
cs.CV2022
Visual Subtitle Feature Enhanced Video Outline Generation
Qi Lv, Ziqiang Cao, Wenrui Xie +10
With the tremendously increasing number of videos, there is a great demand for techniques that help people quickly navigate to the video segments they are interested in. However, c…
cs.CV2022★ 1 cited
Revising Image-Text Retrieval via Multi-Modal Entailment
Xu Yan, Chunhui Ai, Ziqiang Cao +4
An outstanding image-text retrieval model depends on high-quality labeled data. While the builders of existing image-text retrieval datasets strive to ensure that the caption match…
cs.CV2021★ 5 cited
Learning Semantic-Aligned Feature Representation for Text-based Person Search
Shiping Li, Min Cao, Min Zhang
Text-based person search aims to retrieve images of a certain pedestrian by a textual description. The key challenge of this task is to eliminate the inter-modality gap and achieve…