233 citations · 247 across the 5 of their papers we have counts for
6 papers
Genie: Generative Interactive Environments
Jake Bruce, Michael Dennis, Ashley Edwards +22
We introduce Genie, the first generative interactive environment trained in an unsupervised manner from unlabelled Internet videos. The model can be prompted to generate an endless…
RoboTAP: Tracking Arbitrary Points for Few-Shot Visual Imitation
Mel Vecerik, Carl Doersch, Yi Yang +6
For robots to be useful outside labs and specialized factories we need a way to teach them new useful behaviors quickly. Current approaches lack either the generality to onboard ne…
Lossless Adaptation of Pretrained Vision Models For Robotic Manipulation
Mohit Sharma, Claudio Fantacci, Yuxiang Zhou +4
Recent works have shown that large models pretrained on common visual learning tasks can provide useful representations for a wide range of specialized perception problems, as well…
Wish you were here: Hindsight Goal Selection for long-horizon dexterous manipulation
Todor Davchev, Oleg Sushkov, Jean-Baptiste Regli +4
Complex sequential tasks in continuous-control settings often require agents to successfully traverse a set of "narrow passages" in their state space. Solving such tasks with a spa…
SoundNet: Learning Sound Representations from Unlabeled Video
Yusuf Aytar, Carl Vondrick, Antonio Torralba
We learn rich natural sound representations by capitalizing on large amounts of unlabeled sound data collected in the wild. We leverage the natural synchronization between vision a…
How Transferable are CNN-based Features for Age and Gender Classification?
Gökhan Özbulak, Yusuf Aytar, Hazım Kemal Ekenel
Age and gender are complementary soft biometric traits for face recognition. Successful estimation of age and gender from facial images taken under real-world conditions can contri…