most citedSignBERT+: Hand-model-aware Self-supervised Pre-training for Sign Language Understanding

137 citations · 146 across the 3 of their papers we have counts for

collaborators

9 papers

cs.CV2024

Scaling up Multimodal Pre-training for Sign Language Understanding

Wengang Zhou, Weichao Zhao, Hezhen Hu +2

Sign language serves as the primary meaning of communication for the deaf-mute community. Different from spoken language, it commonly conveys information by the collaboration of ma…

cs.CV20241 cited

Expressive Gaussian Human Avatars from Monocular RGB Video

Hezhen Hu, Zhiwen Fan, Tianhao Wu +4

Nuanced expressiveness, particularly through fine-grained hand and facial expressions, is pivotal for enhancing the realism and vitality of digital human representations. In this w…

cs.CV2024

Self-Supervised Representation Learning with Spatial-Temporal Consistency for Sign Language Recognition

Weichao Zhao, Wengang Zhou, Hezhen Hu +2

Recently, there have been efforts to improve the performance in sign language recognition by designing self-supervised learning methods. However, these methods capture limited info…

cs.CV20241 cited

Comp4D: LLM-Guided Compositional 4D Scene Generation

Dejia Xu, Hanwen Liang, Neel P. Bhatt +4

Recent advancements in diffusion models for 2D and 3D content creation have sparked a surge of interest in generating 4D content. However, the scarcity of 3D scene datasets constra…

cs.CV20231 cited

PersonMAE: Person Re-Identification Pre-Training with Masked AutoEncoders

Hezhen Hu, Xiaoyi Dong, Jianmin Bao +4

Pre-training is playing an increasingly important role in learning generic feature representation for Person Re-identification (ReID). We argue that a high-quality ReID representat…

cs.CV20232 cited

Sign Language Translation with Iterative Prototype

Huijie Yao, Wengang Zhou, Hao Feng +3

This paper presents IP-SLT, a simple yet effective framework for sign language translation (SLT). Our IP-SLT adopts a recurrent structure and enhances the semantic representation (…