5 citations · 6 across the 3 of their papers we have counts for
3 papers
cs.CV2024
360+x: A Panoptic Multi-modal Scene Understanding Dataset
Hao Chen, Yuqi Hou, Chenyuan Qu +3
Human perception of the world is shaped by a multitude of viewpoints and modalities. While many existing datasets focus on scene understanding from a certain perspective (e.g. egoc…
cs.CV2024★ 5 cited
SignVTCL: Multi-Modal Continuous Sign Language Recognition Enhanced by Visual-Textual Contrastive Learning
Hao Chen, Jiaze Wang, Ziyu Guo +6
Sign language recognition (SLR) plays a vital role in facilitating communication for the hearing-impaired community. SLR is a weakly supervised task where entire videos are annotat…
cs.CV2023★ 1 cited
Hierarchical Cross-modal Transformer for RGB-D Salient Object Detection
Hao Chen, Feihong Shen
Most of existing RGB-D salient object detection (SOD) methods follow the CNN-based paradigm, which is unable to model long-range dependencies across space and modalities due to the…