3 citations · 3 across the 1 of their papers we have counts for
1 paper
Srikar Appalaraju, Peng Tang, Qi Dong +3
We propose DocFormerv2, a multi-modal transformer for Visual Document Understanding (VDU). The VDU domain entails understanding documents (beyond mere OCR predictions) e.g., extrac…