5 citations · 9 across the 2 of their papers we have counts for
2 papers
cs.CV2023★ 4 cited
DocAligner: Annotating Real-world Photographic Document Images by Simply Taking Pictures
Jiaxin Zhang, Bangdong Chen, Hiuyi Cheng +3
Recently, there has been a growing interest in research concerning document image analysis and recognition in photographic scenarios. However, the lack of labeled datasets for this…
cs.CV2023★ 5 cited
MDoc: A Large-Scale Multi-Format, Multi-Type, Multi-Layout, Multi-Language, Multi-Annotation Category Dataset for Modern Document Layout Analysis
Hiuyi Cheng, Peirong Zhang, Sihang Wu +6
Document layout analysis is a crucial prerequisite for document understanding, including document retrieval and conversion. Most public datasets currently contain only PDF document…