97 citations · 97 across the 2 of their papers we have counts for
4 papers · 1 filter
Retrieving Any Relevant Moments: Benchmark and Models for Generalized Moment Retrieval
Yiming Ding, Siyu Cao, Luyuan Jiao +4
Video Moment Retrieval (VMR) aims to localize temporal segments in videos that correspond to a natural language query, but typically assumes only a single matching moment for each…
Boosting Open-Vocabulary Object Detection by Handling Background Samples
Ruizhe Zeng, Lu Zhang, Xu Yang +1
Open-vocabulary object detection is the task of accurately detecting objects from a candidate vocabulary list that includes both base and novel categories. Currently, numerous open…
Weakly Aligned Feature Fusion for Multimodal Object Detection
Lu Zhang, Zhiyong Liu, Xiangyu Zhu +4
To achieve accurate and robust object detection in the real-world scenario, various forms of images are incorporated, such as color, thermal, and depth. However, multimodal data of…
Weakly Aligned Cross-Modal Learning for Multispectral Pedestrian Detection
Lu Zhang, Xiangyu Zhu, Xiangyu Chen +3
Multispectral pedestrian detection has shown great advantages under poor illumination conditions, since the thermal modality provides complementary information for the color image.…