331 citations · 449 across the 5 of their papers we have counts for
5 papers
Visual Semantic Role Labeling
Saurabh Gupta, Jitendra Malik
In this paper we introduce the problem of Visual Semantic Role Labeling: given an image we want to detect people doing actions and localize the objects of interaction. Classical ap…
Inferring 3D Object Pose in RGB-D Images
Saurabh Gupta, Pablo Arbeláez, Ross Girshick +1
The goal of this work is to replace objects in an RGB-D scene with corresponding 3D models from a library. We approach this problem by first detecting and segmenting object instanc…
Detecting People in Cubist Art
Shiry Ginosar, Daniel Haas, Timothy Brown +1
Although the human visual system is surprisingly robust to extreme distortion when recognizing objects, most evaluations of computer object detection methods focus only on robustne…
Learning Rich Features from RGB-D Images for Object Detection and Segmentation
Saurabh Gupta, Ross Girshick, Pablo Arbeláez +1
In this paper we study the problem of object detection for RGB-D images using semantically rich image and depth features. We propose a new geocentric embedding for depth images tha…
Pixels to Voxels: Modeling Visual Representation in the Human Brain
Pulkit Agrawal, Dustin Stansbury, Jitendra Malik +1
The human brain is adept at solving difficult high-level visual processing problems such as image interpretation and object recognition in natural scenes. Over the past few years n…