3 citations · 5 across the 4 of their papers we have counts for
4 papers · 1 filter
FlashVTG: Feature Layering and Adaptive Score Handling Network for Video Temporal Grounding
Zhuo Cao, Bingqing Zhang, Heming Du +3
Text-guided Video Temporal Grounding (VTG) aims to localize relevant segments in untrimmed videos based on textual descriptions, encompassing two subtasks: Moment Retrieval (MR) an…
Affective Behaviour Analysis via Integrating Multi-Modal Knowledge
Wei Zhang, Feng Qiu, Chen Liu +4
Affective Behavior Analysis aims to facilitate technology emotionally smart, creating a world where devices can understand and react to our emotions as humans do. To comprehensivel…
When 3D Bounding-Box Meets SAM: Point Cloud Instance Segmentation with Weak-and-Noisy Supervision
Qingtao Yu, Heming Du, Chen Liu +1
Learning from bounding-boxes annotations has shown great potential in weakly-supervised 3D point cloud instance segmentation. However, we observed that existing methods would suffe…
RVD: A Handheld Device-Based Fundus Video Dataset for Retinal Vessel Segmentation
MD Wahiduzzaman Khan, Hongwei Sheng, Hu Zhang +11
Retinal vessel segmentation is generally grounded in image-based datasets collected with bench-top devices. The static images naturally lose the dynamic characteristics of retina f…