105 citations · 129 across the 9 of their papers we have counts for
14 papers
Neural 3D Reconstruction in the Wild
Jiaming Sun, Xi Chen, Qianqian Wang +4
We are witnessing an explosion of neural implicit representations in computer vision and graphics. Their applicability has recently expanded beyond tasks such as shape generation a…
Who's Waldo? Linking People Across Text and Images
Claire Yuqing Cui, Apoorv Khandelwal, Yoav Artzi +2
We present a task and benchmark dataset for person-centric visual grounding, the problem of linking between people named in a caption and people pictured in an image. In contrast t…
Towers of Babel: Combining Images, Language, and 3D Geometry for Learning Multimodal Vision
Xiaoshi Wu, Hadar Averbuch-Elor, Jin Sun +1
The abundance and richness of Internet photos of landmarks and cities has led to significant progress in 3D vision over the past two decades, including automated 3D reconstructions…
Extreme Rotation Estimation using Dense Correlation Volumes
Ruojin Cai, Bharath Hariharan, Noah Snavely +1
We present a technique for estimating the relative 3D rotation of an RGB image pair in an extreme setting, where the images have little or no overlap. We observe that, even when im…
An Ethical Highlighter for People-Centric Dataset Creation
Margot Hanley, Apoorv Khandelwal, Hadar Averbuch-Elor +2
Important ethical concerns arising from computer vision datasets of people have been receiving significant attention, and a number of datasets have been withdrawn as a result. To m…
Hidden Footprints: Learning Contextual Walkability from 3D Human Trails
Jin Sun, Hadar Averbuch-Elor, Qianqian Wang +1
Predicting where people can walk in a scene is important for many tasks, including autonomous driving systems and human behavior analysis. Yet learning a computational model for th…