7 citations · 7 across the 3 of their papers we have counts for
6 papers
Slow-Fast Auditory Streams For Audio Recognition
Evangelos Kazakos, Arsha Nagrani, Andrew Zisserman +1
We propose a two-stream convolutional network for audio recognition, that operates on time-frequency spectrogram inputs. Following similar success in visual recognition, we learn S…
The EPIC-KITCHENS Dataset: Collection, Challenges and Baselines
Dima Damen, Hazel Doughty, Giovanni Maria Farinella +8
Since its introduction in 2018, EPIC-KITCHENS has attracted attention as the largest egocentric video benchmark, offering a unique viewpoint on people's interaction with objects, t…
EPIC-Fusion: Audio-Visual Temporal Binding for Egocentric Action Recognition
Evangelos Kazakos, Arsha Nagrani, Andrew Zisserman +1
We focus on multi-modal fusion for egocentric action recognition, and propose a novel architecture for multi-modal temporal-binding, i.e. the combination of modalities within a ran…
Scaling Egocentric Vision: The EPIC-KITCHENS Dataset
Dima Damen, Hazel Doughty, Giovanni Maria Farinella +8
First-person vision is gaining interest as it offers a unique viewpoint on people's interaction with objects, their attention, and even intention. However, progress in this challen…
Human Activity Recognition Using Robust Adaptive Privileged Probabilistic Learning
Michalis Vrigkas, Evangelos Kazakos, Christophoros Nikou +1
In this work, a novel method based on the learning using privileged information (LUPI) paradigm for recognizing complex human activities is proposed that handles missing informatio…
Inferring Human Activities Using Robust Privileged Probabilistic Learning
Michalis Vrigkas, Evangelos Kazakos, Christophoros Nikou +1
Classification models may often suffer from "structure imbalance" between training and testing data that may occur due to the deficient data collection process. This imbalance can…