4 citations · 4 across the 1 of their papers we have counts for
1 paper · 1 filter
Srikar Appalaraju, Bhavan Jasani, Bhargava Urala Kota +2
We present DocFormer -- a multi-modal transformer based architecture for the task of Visual Document Understanding (VDU). VDU is a challenging problem which aims to understand docu…