4 citations · 4 across the 1 of their papers we have counts for
1 paper
Jiaqi Gu, Hyoukjun Kwon, Dilin Wang +6
Vision Transformers (ViTs) have emerged with superior performance on computer vision tasks compared to convolutional neural network (CNN)-based models. However, ViTs are mainly des…