1 paper
Daniel A. P. Oliveira, David Martins de Matos
The computer vision community has seen a shift from convolutional-based to pure transformer architectures for both image and video tasks. Training a transformer from zero for these…