3 citations · 3 across the 1 of their papers we have counts for
4 papers
ViBERTgrid: A Jointly Trained Multi-Modal 2D Document Representation for Key Information Extraction from Documents
Weihong Lin, Qifang Gao, Lei Sun +4
Recent grid-based document representations like BERTgrid allow the simultaneous encoding of the textual and layout information of a document in a 2D feature map so that state-of-th…
ReLaText: Exploiting Visual Relationships for Arbitrary-Shaped Scene Text Detection with Graph Convolutional Networks
Chixiang Ma, Lei Sun, Zhuoyao Zhong +1
We introduce a new arbitrary-shaped text detection approach named ReLaText by formulating text detection as a visual relationship detection problem. To demonstrate the effectivenes…
Mask R-CNN with Pyramid Attention Network for Scene Text Detection
Zhida Huang, Zhuoyao Zhong, Lei Sun +1
In this paper, we present a new Mask R-CNN based text detection approach which can robustly detect multi-oriented and curved text from natural scene images in a unified manner. To…
An Anchor-Free Region Proposal Network for Faster R-CNN based Text Detection Approaches
Zhuoyao Zhong, Lei Sun, Qiang Huo
The anchor mechanism of Faster R-CNN and SSD framework is considered not effective enough to scene text detection, which can be attributed to its IoU based matching criterion betwe…