78 citations · 114 across the 6 of their papers we have counts for
8 papers
Hierarchical Contrast for Unsupervised Skeleton-based Action Representation Learning
Jianfeng Dong, Shengkai Sun, Zhonglin Liu +3
This paper targets unsupervised skeleton-based action representation learning and proposes a new Hierarchical Contrast (HiCo) framework. Different from the existing contrastive-bas…
Modality-Balanced Embedding for Video Retrieval
Xun Wang, Bingqing Ke, Xuanping Li +6
Video search has become the main routine for users to discover videos relevant to a text query on large short-video sharing platforms. During training a query-video bi-encoder mode…
Reading-strategy Inspired Visual Representation Learning for Text-to-Video Retrieval
Jianfeng Dong, Yabing Wang, Xianke Chen +4
This paper aims for the task of text-to-video retrieval, where given a query in the form of a natural-language sentence, it is asked to retrieve videos which are semantically relev…
Deep Dual Consecutive Network for Human Pose Estimation
Zhenguang Liu, Haoming Chen, Runyang Feng +4
Multi-frame human pose estimation in complicated situations is challenging. Although state-of-the-art human joints detectors have demonstrated remarkable results for static images,…
Dual Encoding for Video Retrieval by Text
Jianfeng Dong, Xirong Li, Chaoxi Xu +4
This paper attacks the challenging problem of video retrieval by text. In such a retrieval paradigm, an end user searches for unlabeled videos by ad-hoc queries described exclusive…
Tree-Augmented Cross-Modal Encoding for Complex-Query Video Retrieval
Xun Yang, Jianfeng Dong, Yixin Cao +3
The rapid growth of user-generated videos on the Internet has intensified the need for text-based video retrieval systems. Traditional methods mainly favor the concept-based paradi…