1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.CL2022
Relational Representation Learning in Visually-Rich Documents
Xin Li, Yan Zheng, Yiqing Hu +5
Relational understanding is critical for a number of visually-rich documents (VRDs) understanding tasks. Through multi-modal pre-training, recent studies provide comprehensive cont…
cs.CV2022★ 1 cited
The Devil is in the Frequency: Geminated Gestalt Autoencoder for Self-Supervised Visual Pre-Training
Hao Liu, Xinghua Jiang, Xin Li +3
The self-supervised Masked Image Modeling (MIM) schema, following "mask-and-reconstruct" pipeline of recovering contents from masked image, has recently captured the increasing int…