895 citations · 1k across the 9 of their papers we have counts for
11 papers · 1 filter
PAN++: Towards Efficient and Accurate End-to-End Spotting of Arbitrarily-Shaped Text
Wenhai Wang, Enze Xie, Xiang Li +5
Scene text detection and recognition have been well explored in the past few years. Despite the progress, efficient and accurate end-to-end spotting of arbitrarily-shaped text rema…
Inter-class Discrepancy Alignment for Face Recognition
Jiaheng Liu, Yudong Wu, Yichao Wu +4
The field of face recognition (FR) has witnessed great progress with the surge of deep learning. Existing methods mainly focus on extracting discriminative features, and directly c…
Segmenting Transparent Object in the Wild with Transformer
Enze Xie, Wenjia Wang, Wenhai Wang +4
This work presents a new fine-grained transparent object segmentation dataset, termed Trans10K-v2, extending Trans10K-v1, the first large-scale transparent object segmentation data…
Pyramid Vision Transformer: A Versatile Backbone for Dense Prediction without Convolutions
Wenhai Wang, Enze Xie, Xiang Li +6
Although using convolutional neural networks (CNNs) as backbones achieves great successes in computer vision, this work investigates a simple backbone network useful for many dense…
DocStruct: A Multimodal Method to Extract Hierarchy Structure in Document for General Form Understanding
Zilong Wang, Mingjie Zhan, Xuebo Liu +1
Form understanding depends on both textual contents and organizational structure. Although modern OCR performs well, it is still challenging to realize general form understanding b…
Scene Text Image Super-Resolution in the Wild
Wenjia Wang, Enze Xie, Xuebo Liu +4
Low-resolution text images are often seen in natural scenes such as documents captured by mobile phones. Recognizing low-resolution text images is challenging because they lose det…