16 citations · 26 across the 4 of their papers we have counts for
3 papers · 1 filter
DocAligner: Annotating Real-world Photographic Document Images by Simply Taking Pictures
Jiaxin Zhang, Bangdong Chen, Hiuyi Cheng +3
Recently, there has been a growing interest in research concerning document image analysis and recognition in photographic scenarios. However, the lack of labeled datasets for this…
MDoc: A Large-Scale Multi-Format, Multi-Type, Multi-Layout, Multi-Language, Multi-Annotation Category Dataset for Modern Document Layout Analysis
Hiuyi Cheng, Peirong Zhang, Sihang Wu +6
Document layout analysis is a crucial prerequisite for document understanding, including document retrieval and conversion. Most public datasets currently contain only PDF document…
Don't Forget Me: Accurate Background Recovery for Text Removal via Modeling Local-Global Context
Chongyu Liu, Lianwen Jin, Yuliang Liu +4
Text removal has attracted increasingly attention due to its various applications on privacy protection, document restoration, and text editing. It has shown significant progress w…