1 paper · 1 filter
Jie Ma, Yalong Bai, Bineng Zhong +3
Vision Transformer (ViT) has become a leading tool in various computer vision tasks, owing to its unique self-attention mechanism that learns visual representations explicitly thro…