activity
20172022
most citedTouch and Go: Learning from Human-Collected Vision and Touch

20 citations · 20 across the 4 of their papers we have counts for

collaborators

15 papers

cs.CV202220 cited

Touch and Go: Learning from Human-Collected Vision and Touch

Fengyu Yang, Chenyang Ma, Jiacheng Zhang +3

The ability to associate touch with sight is essential for tasks that require physically interacting with objects in the world. We propose a dataset with paired visual and tactile…

cs.CV2022

Mix and Localize: Localizing Sound Sources in Mixtures

Xixi Hu, Ziyang Chen, Andrew Owens

We present a method for simultaneously localizing multiple sound sources within a visual scene. This task requires a model to both group a sound mixture into individual sources, an…

cs.CV2022

Learning Visual Styles from Audio-Visual Associations

Tingle Li, Yichen Liu, Andrew Owens +1

From the patter of rain to the crunch of snow, the sounds we hear often convey the visual textures that appear within a scene. In this paper, we present a method for learning visua…

cs.CV2021

Strumming to the Beat: Audio-Conditioned Contrastive Video Textures

Medhini Narasimhan, Shiry Ginosar, Andrew Owens +2

We introduce a non-parametric approach for infinite video texture synthesis using a representation learned via contrastive learning. We take inspiration from Video Textures, which…

cs.CV2021

Planar Surface Reconstruction from Sparse Views

Linyi Jin, Shengyi Qian, Andrew Owens +1

The paper studies planar surface reconstruction of indoor scenes from two views with unknown camera poses. While prior approaches have successfully created object-centric reconstru…

cs.CV2020

Self-Supervised Learning of Audio-Visual Objects from Video

Triantafyllos Afouras, Andrew Owens, Joon Son Chung +1

Our objective is to transform a video into a set of discrete audio-visual objects using self-supervised learning. To this end, we introduce a model that uses attention to localize…