209 citations · 322 across the 5 of their papers we have counts for
5 papers
Attentional Pooling for Action Recognition
Rohit Girdhar, Deva Ramanan
We introduce a simple yet surprisingly powerful model to incorporate attention in action recognition and human object interaction tasks. Our proposed attention module can be traine…
Tinkering Under the Hood: Interactive Zero-Shot Learning with Net Surgery
Vivek Krishnan, Deva Ramanan
We consider the task of visual net surgery, in which a CNN can be reconfigured without extra data to recognize novel concepts that may be omitted from the training set. While most…
PixelNet: Towards a General Pixel-level Architecture
Aayush Bansal, Xinlei Chen, Bryan Russell +2
We explore architectures for general pixel-level prediction problems, from low-level edge detection to mid-level surface normal estimation to high-level semantic segmentation. Conv…
3D Hand Pose Detection in Egocentric RGB-D Images
Gregory Rogez, James S. Supancic, Maryam Khademi +2
We focus on the task of everyday hand pose estimation from egocentric viewpoints. For this task, we show that depth sensors are particularly informative for extracting near-field i…
Egocentric Pose Recognition in Four Lines of Code
Gregory Rogez, James S. Supancic, Deva Ramanan
We tackle the problem of estimating the 3D pose of an individual's upper limbs (arms+hands) from a chest mounted depth-camera. Importantly, we consider pose estimation during every…