77 citations · 135 across the 8 of their papers we have counts for
19 papers
Learn what matters: cross-domain imitation learning with task-relevant embeddings
Tim Franzmeyer, Philip H. S. Torr, João F. Henriques
We study how an autonomous agent learns to perform a task from demonstrations in a different domain, such as a different environment or different agent. Such cross-domain imitation…
A 23 MW data centre is all you need
Samuel Albanie, Dylan Campbell, João F. Henriques
The field of machine learning has achieved striking progress in recent years, witnessing breakthrough results on language modelling, protein folding and nitpickingly fine-grained d…
Audio Retrieval with Natural Language Queries: A Benchmark Study
A. Sophia Koepke, Andreea-Maria Oncescu, João F. Henriques +2
The objectives of this work are cross-modal text-audio and audio-text retrieval, in which the goal is to retrieve the audio content from a pool of candidates that best matches a gi…
Keeping Your Eye on the Ball: Trajectory Attention in Video Transformers
Mandela Patrick, Dylan Campbell, Yuki M. Asano +5
In video transformers, the time dimension is often treated in the same way as the two spatial dimensions. However, in a scene where objects or the camera may move, a physical point…
Moving SLAM: Fully Unsupervised Deep Learning in Non-Rigid Scenes
Dan Xu, Andrea Vedaldi, Joao F. Henriques
We propose a method to train deep networks to decompose videos into 3D geometry (camera and depth), moving objects, and their motions, with no supervision. We build on the idea of…
Audio Retrieval with Natural Language Queries
Andreea-Maria Oncescu, A. Sophia Koepke, João F. Henriques +2
We consider the task of retrieving audio using free-form natural language queries. To study this problem, which has received limited attention in the existing literature, we introd…