Publications (12)
Domain Adaptive Person Search
Junjie Li, Yichao Yan, Guanshuo Wang +3
Person search is a challenging task which aims to achieve joint pedestrian detection and person re-identification (ReID). Previous works have made significant advances under fully…
MS-DETR: Towards Effective Video Moment Retrieval and Highlight Detection by Joint Motion-Semantic Learning
Hongxu Ma, Guanshuo Wang, Fufu Yu +2
Video Moment Retrieval (MR) and Highlight Detection (HD) aim to pinpoint specific moments and assess clip-wise relevance based on the text query. While DETR-based joint frameworks…
Exploiting the Textual Potential from Vision-Language Pre-training for Text-based Person Search
Guanshuo Wang, Fufu Yu, Junjie Li +2
Text-based Person Search (TPS), is targeted on retrieving pedestrians to match text descriptions instead of query images. Recent Vision-Language Pre-training (VLP) models can bring…
Rethinking Clothes Changing Person ReID: Conflicts, Synthesis, and Optimization
Junjie Li, Guanshuo Wang, Fufu Yu +6
Clothes-changing person re-identification (CC-ReID) aims to retrieve images of the same person wearing different outfits. Mainstream researches focus on designing advanced model st…
D2Pruner: Debiased Importance and Structural Diversity for MLLM Token Pruning
Evelyn Zhang, Fufu Yu, Aoqi Wu +5
Processing long visual token sequences poses a significant computational burden on Multimodal Large Language Models (MLLMs). While token pruning offers a path to acceleration, we f…
SoccerNet 2023 Challenges Results
Anthony Cioppa, Silvio Giancola, Vladimir Somers +99
The SoccerNet 2023 challenges were the third annual video understanding challenges organized by the SoccerNet team. For this third edition, the challenges were composed of seven vi…