1 paper · 1 filter
Seungeun Oh, Sihun Baek, Jihong Park +5
In computer vision, the vision transformer (ViT) has increasingly superseded the convolutional neural network (CNN) for improved accuracy and robustness. However, ViT's large model…