9 citations · 10 across the 6 of their papers we have counts for
6 papers
Dynamic Relation Transformer for Contextual Text Block Detection
Jiawei Wang, Shunchi Zhang, Kai Hu +4
Contextual Text Block Detection (CTBD) is the task of identifying coherent text blocks within the complexity of natural scenes. Previous methodologies have treated CTBD as either a…
UniVIE: A Unified Label Space Approach to Visual Information Extraction from Form-like Documents
Kai Hu, Jiawei Wang, Weihong Lin +3
Existing methods for Visual Information Extraction (VIE) from form-like documents typically fragment the process into separate subtasks, such as key information extraction, key-val…
Improving Handwritten OCR with Training Samples Generated by Glyph Conditional Denoising Diffusion Probabilistic Model
Haisong Ding, Bozhi Luan, Dongnan Gui +2
Constructing a highly accurate handwritten OCR system requires large amounts of representative training data, which is both time-consuming and expensive to collect. To mitigate the…
Zero-shot Generation of Training Data with Denoising Diffusion Probabilistic Model for Handwritten Chinese Character Recognition
Dongnan Gui, Kai Chen, Haisong Ding +1
There are more than 80,000 character categories in Chinese while most of them are rarely used. To build a high performance handwritten Chinese character recognition (HCCR) system s…
A Question-Answering Approach to Key Value Pair Extraction from Form-like Document Images
Kai Hu, Zhuoyuan Wu, Zhuoyao Zhong +3
In this paper, we present a new question-answering (QA) based key-value pair extraction approach, called KVPFormer, to robustly extracting key-value relationships between entities…
TSRFormer: Table Structure Recognition with Transformers
Weihong Lin, Zheng Sun, Chixiang Ma +4
We present a new table structure recognition (TSR) approach, called TSRFormer, to robustly recognizing the structures of complex tables with geometrical distortions from various ta…