36 citations · 47 across the 3 of their papers we have counts for
3 papers
cs.CV2022★ 36 cited
Uniform Masking: Enabling MAE Pre-training for Pyramid-based Vision Transformers with Locality
Xiang Li, Wenhai Wang, Lingfeng Yang +1
Masked AutoEncoder (MAE) has recently led the trends of visual self-supervision area by an elegant asymmetric encoder-decoder design, which significantly optimizes both the pre-tra…
cs.CV2022★ 10 cited
RecursiveMix: Mixed Learning with History
Lingfeng Yang, Xiang Li, Borui Zhao +2
Mix-based augmentation has been proven fundamental to the generalization of deep vision models. However, current augmentations only mix samples at the current data batch during tra…
cs.CV2022★ 1 cited
Dynamic MLP for Fine-Grained Image Classification by Leveraging Geographical and Temporal Information
Lingfeng Yang, Xiang Li, Renjie Song +5
Fine-grained image classification is a challenging computer vision task where various species share similar visual appearances, resulting in misclassification if merely based on vi…