20 citations · 51 across the 22 of their papers we have counts for
Showing 2022Show all
2 papers · 1 filter
cs.CV2022★ 2 cited
Multimodal Tree Decoder for Table of Contents Extraction in Document Images
Pengfei Hu, Zhenrong Zhang, Jianshu Zhang +2
Table of contents (ToC) extraction aims to extract headings of different levels in documents to better understand the outline of the contents, which can be widely used for document…
cs.CV2022★ 1 cited
Multimodal Pre-training Based on Graph Attention Network for Document Understanding
Zhenrong Zhang, Jiefeng Ma, Jun Du +2
Document intelligence as a relatively new research topic supports many business applications. Its main task is to automatically read, understand, and analyze documents. However, du…