18 citations · 25 across the 7 of their papers we have counts for
Showing 2023 · cs.CVShow all
3 papers · 2 filters
cs.CV2023★ 1 cited
MataDoc: Margin and Text Aware Document Dewarping for Arbitrary Boundary
Beiya Dai, Xing li, Qunyi Xie +5
Document dewarping from a distorted camera-captured image is of great value for OCR and document understanding. The document boundary plays an important role which is more evident…
cs.CV2023
Fast-StrucTexT: An Efficient Hourglass Transformer with Modality-guided Dynamic Token Merge for Document Understanding
Mingliang Zhai, Yulin Li, Xiameng Qin +6
Transformers achieve promising performance in document understanding because of their high effectiveness and still suffer from quadratic computational complexity dependency on the…
cs.CV2023★ 18 cited
StrucTexTv2: Masked Visual-Textual Prediction for Document Image Pre-training
Yuechen Yu, Yulin Li, Chengquan Zhang +7
In this paper, we present StrucTexTv2, an effective document image pre-training framework, by performing masked visual-textual prediction. It consists of two self-supervised pre-tr…