3 citations · 5 across the 5 of their papers we have counts for
5 papers
TokenBinder: Text-Video Retrieval with One-to-Many Alignment Paradigm
Bingqing Zhang, Zhuo Cao, Heming Du +4
Text-Video Retrieval (TVR) methods typically match query-candidate pairs by aligning text and video features in coarse-grained, fine-grained, or combined (coarse-to-fine) manners.…
Affective Behaviour Analysis via Integrating Multi-Modal Knowledge
Wei Zhang, Feng Qiu, Chen Liu +4
Affective Behavior Analysis aims to facilitate technology emotionally smart, creating a world where devices can understand and react to our emotions as humans do. To comprehensivel…
Divide and Ensemble: Progressively Learning for the Unknown
Hu Zhang, Xin Shen, Heming Du +10
In the wheat nutrient deficiencies classification challenge, we present the DividE and EnseMble (DEEM) method for progressive test data predictions. We find that (1) test images ar…
When 3D Bounding-Box Meets SAM: Point Cloud Instance Segmentation with Weak-and-Noisy Supervision
Qingtao Yu, Heming Du, Chen Liu +1
Learning from bounding-boxes annotations has shown great potential in weakly-supervised 3D point cloud instance segmentation. However, we observed that existing methods would suffe…
RVD: A Handheld Device-Based Fundus Video Dataset for Retinal Vessel Segmentation
MD Wahiduzzaman Khan, Hongwei Sheng, Hu Zhang +11
Retinal vessel segmentation is generally grounded in image-based datasets collected with bench-top devices. The static images naturally lose the dynamic characteristics of retina f…