19 citations · 40 across the 5 of their papers we have counts for
5 papers
Fine-grained Semantic Alignment Network for Weakly Supervised Temporal Language Grounding
Yuechen Wang, Wengang Zhou, Houqiang Li
Temporal language grounding (TLG) aims to localize a video segment in an untrimmed video based on a natural language description. To alleviate the expensive cost of manual annotati…
Geometric Representation Learning for Document Image Rectification
Hao Feng, Wengang Zhou, Jiajun Deng +2
In document image rectification, there exist rich geometric constraints between the distorted image and the ground truth one. However, such geometric constraints are largely ignore…
SignBERT: Pre-Training of Hand-Model-Aware Representation for Sign Language Recognition
Hezhen Hu, Weichao Zhao, Wengang Zhou +2
Hand gesture serves as a critical role in sign language. Current deep-learning-based sign language recognition (SLR) methods may suffer insufficient interpretability and overfittin…
Weakly Supervised Temporal Adjacent Network for Language Grounding
Yuechen Wang, Jiajun Deng, Wengang Zhou +1
Temporal language grounding (TLG) is a fundamental and challenging problem for vision and language understanding. Existing methods mainly focus on fully supervised setting with tem…
Pre-training Text Representations as Meta Learning
Shangwen Lv, Yuechen Wang, Daya Guo +10
Pre-training text representations has recently been shown to significantly improve the state-of-the-art in many natural language processing tasks. The central goal of pre-training…