3 citations · 3 across the 1 of their papers we have counts for
1 paper
Yanhong Fei, Yingjie Liu, Xian Wei +1
Inspired by the tremendous success of the self-attention mechanism in natural language processing, the Vision Transformer (ViT) creatively applies it to image patch sequences and a…