4 citations · 4 across the 1 of their papers we have counts for
1 paper
Benjia Zhou, Pichao Wang, Jun Wan +2
Vision Transformers (ViTs) have shown promising performance compared with Convolutional Neural Networks (CNNs), but the training of ViTs is much harder than CNNs. In this paper, we…