5 citations · 6 across the 4 of their papers we have counts for
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2022
Visual Subtitle Feature Enhanced Video Outline Generation
Qi Lv, Ziqiang Cao, Wenrui Xie +10
With the tremendously increasing number of videos, there is a great demand for techniques that help people quickly navigate to the video segments they are interested in. However, c…
cs.CV2022★ 1 cited
Revising Image-Text Retrieval via Multi-Modal Entailment
Xu Yan, Chunhui Ai, Ziqiang Cao +4
An outstanding image-text retrieval model depends on high-quality labeled data. While the builders of existing image-text retrieval datasets strive to ensure that the caption match…
cs.CV2021★ 5 cited
Learning Semantic-Aligned Feature Representation for Text-based Person Search
Shiping Li, Min Cao, Min Zhang
Text-based person search aims to retrieve images of a certain pedestrian by a textual description. The key challenge of this task is to eliminate the inter-modality gap and achieve…