7 citations · 7 across the 2 of their papers we have counts for
7 papers
On Semantic Similarity in Video Retrieval
Michael Wray, Hazel Doughty, Dima Damen
Current video retrieval efforts all found their evaluation on an instance-based assumption, that only a single caption is relevant to a query video and vice versa. We demonstrate t…
Supervision Levels Scale (SLS)
Dima Damen, Michael Wray
We propose a three-dimensional discrete and incremental scale to encode a method's level of supervision - i.e. the data and labels used when training a model to achieve a given per…
The EPIC-KITCHENS Dataset: Collection, Challenges and Baselines
Dima Damen, Hazel Doughty, Giovanni Maria Farinella +8
Since its introduction in 2018, EPIC-KITCHENS has attracted attention as the largest egocentric video benchmark, offering a unique viewpoint on people's interaction with objects, t…
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings
Michael Wray, Diane Larlus, Gabriela Csurka +1
We address the problem of cross-modal fine-grained action retrieval between text and video. Cross-modal retrieval is commonly achieved through learning a shared embedding space, th…
Learning Visual Actions Using Multiple Verb-Only Labels
Michael Wray, Dima Damen
This work introduces verb-only representations for both recognition and retrieval of visual actions, in video. Current methods neglect legitimate semantic ambiguities between verbs…
Towards an Unequivocal Representation of Actions
Michael Wray, Davide Moltisanti, Dima Damen
This work introduces verb-only representations for actions and interactions; the problem of describing similar motions (e.g. 'open door', 'open cupboard'), and distinguish differin…