5 citations · 6 across the 2 of their papers we have counts for
2 papers
cs.CV2023★ 5 cited
MDoc: A Large-Scale Multi-Format, Multi-Type, Multi-Layout, Multi-Language, Multi-Annotation Category Dataset for Modern Document Layout Analysis
Hiuyi Cheng, Peirong Zhang, Sihang Wu +6
Document layout analysis is a crucial prerequisite for document understanding, including document retrieval and conversion. Most public datasets currently contain only PDF document…
cs.CV2023★ 1 cited
Improving Table Structure Recognition with Visual-Alignment Sequential Coordinate Modeling
Yongshuai Huang, Ning Lu, Dapeng Chen +5
Table structure recognition aims to extract the logical and physical structure of unstructured table images into a machine-readable format. The latest end-to-end image-to-text appr…