27 citations · 50 across the 11 of their papers we have counts for
1 paper · 2 filters
Peizhao Li, Jiuxiang Gu, Jason Kuen +5
We propose SelfDoc, a task-agnostic pre-training framework for document image understanding. Because documents are multimodal and are intended for sequential reading, our framework…