18 citations · 18 across the 1 of their papers we have counts for
1 paper
Yuechen Yu, Yulin Li, Chengquan Zhang +7
In this paper, we present StrucTexTv2, an effective document image pre-training framework, by performing masked visual-textual prediction. It consists of two self-supervised pre-tr…