4 papers
Class Feature Pyramids for Video Explanation
Alexandros Stergiou, Georgios Kapidis, Grigorios Kalliatakis +3
Deep convolutional networks are widely used in video action recognition. 3D convolutions are one prominent approach to deal with the additional time dimension. While 3D convolution…
Multitask Learning to Improve Egocentric Action Recognition
Georgios Kapidis, Ronald Poppe, Elsbeth van Dam +2
In this work we employ multitask learning to capitalize on the structure that exists in related supervised tasks to train complex neural networks. It allows training a network for…
Egocentric Hand Track and Object-based Human Action Recognition
Georgios Kapidis, Ronald Poppe, Elsbeth van Dam +2
Egocentric vision is an emerging field of computer vision that is characterized by the acquisition of images and video from the first person perspective. In this paper we address t…
Saliency Tubes: Visual Explanations for Spatio-Temporal Convolutions
Alexandros Stergiou, Georgios Kapidis, Grigorios Kalliatakis +3
Deep learning approaches have been established as the main methodology for video classification and recognition. Recently, 3-dimensional convolutions have been used to achieve stat…