35 citations · 53 across the 5 of their papers we have counts for
10 papers
Masked Contrastive Pre-Training for Efficient Video-Text Retrieval
Fangxun Shu, Biaolong Chen, Yue Liao +6
We present a simple yet effective end-to-end Video-language Pre-training (VidLP) framework, Masked Contrastive Video-language Pretraining (MAC), for video-text retrieval tasks. Our…
GEN-VLKT: Simplify Association and Enhance Interaction Understanding for HOI Detection
Yue Liao, Aixi Zhang, Miao Lu +3
The task of Human-Object Interaction~(HOI) detection could be divided into two core problems, i.e., human-object association and interaction understanding. In this paper, we reveal…
TR-MOT: Multi-Object Tracking by Reference
Mingfei Chen, Yue Liao, Si Liu +2
Multi-object Tracking (MOT) generally can be split into two sub-tasks, i.e., detection and association. Many previous methods follow the tracking by detection paradigm, which first…
Mining the Benefits of Two-stage and One-stage HOI Detection
Aixi Zhang, Yue Liao, Si Liu +4
Two-stage methods have dominated Human-Object Interaction (HOI) detection for several years. Recently, one-stage HOI detection methods have become popular. In this paper, we aim to…
Human-centric Relation Segmentation: Dataset and Solution
Si Liu, Zitian Wang, Yulu Gao +5
Vision and language understanding techniques have achieved remarkable progress, but currently it is still difficult to well handle problems involving very fine-grained details. For…
Reformulating HOI Detection as Adaptive Set Prediction
Mingfei Chen, Yue Liao, Si Liu +3
Determining which image regions to concentrate on is critical for Human-Object Interaction (HOI) detection. Conventional HOI detectors focus on either detected human and object pai…