8 citations · 9 across the 2 of their papers we have counts for
2 papers
cs.CV2022★ 1 cited
Students taught by multimodal teachers are superior action recognizers
Gorjan Radevski, Dusan Grujicic, Matthew Blaschko +2
The focal point of egocentric video understanding is modelling hand-object interactions. Standard models -- CNNs, Vision Transformers, etc. -- which receive RGB frames as input per…
cs.CV2021★ 8 cited
Revisiting spatio-temporal layouts for compositional action recognition
Gorjan Radevski, Marie-Francine Moens, Tinne Tuytelaars
Recognizing human actions is fundamentally a spatio-temporal reasoning problem, and should be, at least to some extent, invariant to the appearance of the human and the objects inv…