2 citations · 3 across the 4 of their papers we have counts for
4 papers
DocTrack: A Visually-Rich Document Dataset Really Aligned with Human Eye Movement for Machine Reading
Hao Wang, Qingxuan Wang, Yue Li +3
The use of visually-rich documents (VRDs) in various fields has created a demand for Document AI models that can read and comprehend documents like humans, which requires the overc…
UGC: Unified GAN Compression for Efficient Image-to-Image Translation
Yuxi Ren, Jie Wu, Peng Zhang +6
Recent years have witnessed the prevailing progress of Generative Adversarial Networks (GANs) in image-to-image translation. However, the success of these GAN models hinges on pond…
M2HGCL: Multi-Scale Meta-Path Integrated Heterogeneous Graph Contrastive Learning
Yuanyuan Guo, Yu Xia, Rui Wang +3
Inspired by the successful application of contrastive learning on graphs, researchers attempt to impose graph contrastive learning approaches on heterogeneous information networks.…
DLIP: Distilling Language-Image Pre-training
Huafeng Kuang, Jie Wu, Xiawu Zheng +5
Vision-Language Pre-training (VLP) shows remarkable progress with the assistance of extremely heavy parameters, which challenges deployment in real applications. Knowledge distilla…