7 citations · 8 across the 2 of their papers we have counts for
5 papers
An Action Is Worth Multiple Words: Handling Ambiguity in Action Recognition
Kiyoon Kim, Davide Moltisanti, Oisin Mac Aodha +1
Precisely naming the action depicted in a video can be a challenging and oftentimes ambiguous task. In contrast to object instances represented as nouns (e.g. dog, cat, chair, etc.…
The EPIC-KITCHENS Dataset: Collection, Challenges and Baselines
Dima Damen, Hazel Doughty, Giovanni Maria Farinella +8
Since its introduction in 2018, EPIC-KITCHENS has attracted attention as the largest egocentric video benchmark, offering a unique viewpoint on people's interaction with objects, t…
Action Recognition from Single Timestamp Supervision in Untrimmed Videos
Davide Moltisanti, Sanja Fidler, Dima Damen
Recognising actions in videos relies on labelled supervision during training, typically the start and end times of each action instance. This supervision is not only subjective, bu…
Towards an Unequivocal Representation of Actions
Michael Wray, Davide Moltisanti, Dima Damen
This work introduces verb-only representations for actions and interactions; the problem of describing similar motions (e.g. 'open door', 'open cupboard'), and distinguish differin…
Scaling Egocentric Vision: The EPIC-KITCHENS Dataset
Dima Damen, Hazel Doughty, Giovanni Maria Farinella +8
First-person vision is gaining interest as it offers a unique viewpoint on people's interaction with objects, their attention, and even intention. However, progress in this challen…