331 citations · 769 across the 7 of their papers we have counts for
9 papers
Unifying Map and Landmark Based Representations for Visual Navigation
Saurabh Gupta, David Fouhey, Sergey Levine +1
This works presents a formulation for visual navigation that unifies map based spatial reasoning and path planning, with landmark based robust plan execution in noisy environments.…
From Lifestyle Vlogs to Everyday Interactions
David F. Fouhey, Wei-cheng Kuo, Alexei A. Efros +1
A major stumbling block to progress in understanding basic human interactions, such as getting out of bed or opening a refrigerator, is lack of good training data. Most past effort…
Large-Scale 3D Shape Reconstruction and Segmentation from ShapeNet Core55
Li Yi, Lin Shao, Manolis Savva +47
We introduce a large-scale 3D shape understanding benchmark using data and annotation from ShapeNet 3D object database. The benchmark consists of two tasks: part-level segmentation…
Generic 3D Representation via Pose Estimation and Matching
Amir R. Zamir, Tilman Wekel, Pulkit Argrawal +3
Though a large body of computer vision research has investigated developing generic semantic representations, efforts towards developing a similar representation for 3D has been li…
Learning a Multi-View Stereo Machine
Abhishek Kar, Christian Häne, Jitendra Malik
We present a learnt system for multi-view stereopsis. In contrast to recent learning based methods for 3D reconstruction, we leverage the underlying 3D geometry of the problem thro…
Combining Self-Supervised Learning and Imitation for Vision-Based Rope Manipulation
Ashvin Nair, Dian Chen, Pulkit Agrawal +4
Manipulation of deformable objects, such as ropes and cloth, is an important but challenging problem in robotics. We present a learning-based system where a robot takes as input a…