21 citations · 27 across the 5 of their papers we have counts for
5 papers · 1 filter
Revisiting Vision Transformer from the View of Path Ensemble
Shuning Chang, Pichao Wang, Hao Luo +2
Vision Transformers (ViTs) are normally regarded as a stack of transformer layers. In this work, we propose a novel view of ViTs showing that they can be seen as ensemble networks…
Effective Vision Transformer Training: A Data-Centric Perspective
Benjia Zhou, Pichao Wang, Jun Wan +2
Vision Transformers (ViTs) have shown promising performance compared with Convolutional Neural Networks (CNNs), but the training of ViTs is much harder than CNNs. In this paper, we…
Image-to-Video Re-Identification via Mutual Discriminative Knowledge Transfer
Pichao Wang, Fan Wang, Hao Li
The gap in representations between image and video makes Image-to-Video Re-identification (I2V Re-ID) challenging, and recent works formulate this problem as a knowledge distillati…
2nd Place Solution to Google Landmark Retrieval 2021
Zhang Yuqi, Xu Xianzhe, Chen Weihua +4
This paper presents the 2nd place solution to the Google Landmark Retrieval 2021 Competition on Kaggle. The solution is based on a baseline with training tricks from person re-iden…
Exploring the Quality of GAN Generated Images for Person Re-Identification
Yiqi Jiang, Weihua Chen, Xiuyu Sun +3
Recently, GAN based method has demonstrated strong effectiveness in generating augmentation data for person re-identification (ReID), on account of its ability to bridge the gap be…