1 paper
Nhat Thanh Tran, Fanghui Xue, Shuai Zhang +4
We present for video understanding (classification) a split space-time attention model, VideoSEMA, consisting of a scalable and efficient Mamba-like attention (SEMA) block in space…