37 citations · 45 across the 3 of their papers we have counts for
1 paper · 1 filter
Renjun Xu, Kaifan Yang, Ke Liu +1
Vision Transformer (ViT) has achieved remarkable performance in computer vision. However, positional encoding in ViT makes it substantially difficult to learn the intrinsic equivar…