1 paper
Haizhao Jing, Liuwei Wan, Xizhe Xue +2
Recently, the Vision Transformer (ViT) model has replaced the classical Convolutional Neural Network (ConvNet) in various computer vision tasks due to its superior performance. Eve…