6 citations · 6 across the 1 of their papers we have counts for
2 papers
cs.CV2021★ 6 cited
Relational Self-Attention: What's Missing in Attention for Video Understanding
Manjin Kim, Heeseung Kwon, Chunyu Wang +2
Convolution has been arguably the most important feature transform for modern neural networks, leading to the advance of deep learning. Recent emergence of Transformer networks, wh…
cs.CV2021
Learning Self-Similarity in Space and Time as Generalized Motion for Video Action Recognition
Heeseung Kwon, Manjin Kim, Suha Kwak +1
Spatio-temporal convolution often fails to learn motion dynamics in videos and thus an effective motion representation is required for video understanding in the wild. In this pape…