13 citations · 18 across the 6 of their papers we have counts for
1 paper · 1 filter
Hong Zhou, Rui Zhang, Peifeng Lai +4
Nowadays, Vision Transformer (ViT) is widely utilized in various computer vision tasks, owing to its unique self-attention mechanism. However, the model architecture of ViT is comp…