2 papers
cs.CV2024
Cross-Domain Document Layout Analysis Using Document Style Guide
Xingjiao Wu, Luwei Xiao, Xiangcheng Du +5
The document layout analysis (DLA) aims to decompose document images into high-level semantic areas (i.e., figures, tables, texts, and background). Creating a DLA framework with st…
cs.CV2024
Aggregated Text Transformer for Scene Text Detection
Zhao Zhou, Xiangcheng Du, Yingbin Zheng +1
This paper explores the multi-scale aggregation strategy for scene text detection in natural images. We present the Aggregated Text TRansformer(ATTR), which is designed to represen…