1 paper · 1 filter
Anurag Arnab, Mostafa Dehghani, Georg Heigold +3
We present pure-transformer based models for video classification, drawing upon the recent success of such models in image classification. Our model extracts spatio-temporal tokens…