4 citations · 8 across the 4 of their papers we have counts for
4 papers
Vision Grid Transformer for Document Layout Analysis
Cheng Da, Chuwei Luo, Qi Zheng +1
Document pre-trained models and grid-based models have proven to be very effective on various tasks in Document AI. However, for the document layout analysis (DLA) task, existing d…
LISTER: Neighbor Decoding for Length-Insensitive Scene Text Recognition
Changxu Cheng, Peng Wang, Cheng Da +2
The diversity in length constitutes a significant characteristic of text. Due to the long-tail distribution of text lengths, most existing methods for scene text recognition (STR)…
Multi-Granularity Prediction with Learnable Fusion for Scene Text Recognition
Cheng Da, Peng Wang, Cong Yao
Due to the enormous technical challenges and wide range of applications, scene text recognition (STR) has been an active research topic in computer vision for years. To tackle this…
Levenshtein OCR
Cheng Da, Peng Wang, Cong Yao
A novel scene text recognizer based on Vision-Language Transformer (VLT) is presented. Inspired by Levenshtein Transformer in the area of NLP, the proposed method (named Levenshtei…