5 citations · 7 across the 3 of their papers we have counts for
3 papers
cs.CV2024
Scaling up Multimodal Pre-training for Sign Language Understanding
Wengang Zhou, Weichao Zhao, Hezhen Hu +2
Sign language serves as the primary meaning of communication for the deaf-mute community. Different from spoken language, it commonly conveys information by the collaboration of ma…
cs.CV2024★ 5 cited
Complete Instances Mining for Weakly Supervised Instance Segmentation
Zecheng Li, Zening Zeng, Yuqi Liang +1
Weakly supervised instance segmentation (WSIS) using only image-level labels is a challenging task due to the difficulty of aligning coarse annotations with the finer task. However…
cs.CV2022★ 2 cited
MDS-Net: A Multi-scale Depth Stratification Based Monocular 3D Object Detection Algorithm
Zhouzhen Xie, Yuying Song, Jingxuan Wu +3
Monocular 3D object detection is very challenging in autonomous driving due to the lack of depth information. This paper proposes a one-stage monocular 3D object detection algorith…