1 citations · 1 across the 1 of their papers we have counts for
1 paper · 1 filter
Seungeun Oh, Sihun Baek, Jihong Park +5
In computer vision, the vision transformer (ViT) has increasingly superseded the convolutional neural network (CNN) for improved accuracy and robustness. However, ViT's large model…