1k citations · 1.3k across the 8 of their papers we have counts for
1 paper · 1 filter
Haiping Wu, Bin Xiao, Noel Codella +4
We present in this paper a new architecture, named Convolutional vision Transformer (CvT), that improves Vision Transformer (ViT) in performance and efficiency by introducing convo…