1 paper
Ali Salamatian, Anthony Fuller, Pritam Sarkar +3
Transformers dominate video recognition. They split videos into tokens, and processing them has expensive superlinear computational cost. Yet videos are filled with redundancy, so…