12 citations · 20 across the 2 of their papers we have counts for
3 papers
cs.CV2022★ 12 cited
HiViT: Hierarchical Vision Transformer Meets Masked Image Modeling
Xiaosong Zhang, Yunjie Tian, Wei Huang +4
Recently, masked image modeling (MIM) has offered a new methodology of self-supervised pre-training of vision transformers. A key idea of efficient implementation is to discard the…
cs.CV2021★ 8 cited
Long-tailed Distribution Adaptation
Zhiliang Peng, Wei Huang, Zonghao Guo +3
Recognizing images with long-tailed distributions remains a challenging problem while there lacks an interpretable mechanism to solve this problem. In this study, we formulate Long…
cs.CV2021
Conformer: Local Features Coupling Global Representations for Visual Recognition
Zhiliang Peng, Wei Huang, Shanzhi Gu +4
Within Convolutional Neural Network (CNN), the convolution operations are good at extracting local features but experience difficulty to capture global representations. Within visu…