151 citations · 561 across the 38 of their papers we have counts for
16 papers · 1 filter
The Curious Layperson: Fine-Grained Image Recognition without Expert Labels
Subhabrata Choudhury, Iro Laina, Christian Rupprecht +1
Most of us are not experts in specific fields, such as ornithology. Nonetheless, we do have general image and language understanding capabilities that we use to match what we see t…
NeuralDiff: Segmenting 3D objects that move in egocentric videos
Vadim Tschernezki, Diane Larlus, Andrea Vedaldi
Given a raw video sequence taken from a freely-moving camera, we study the problem of decomposing the observed 3D scene into a static background and a dynamic foreground containing…
Lifting 2D Object Locations to 3D by Discounting LiDAR Outliers across Objects and Views
Robert McCraith, Eldar Insafutdinov, Lukas Neumann +1
We present a system for automatic converting of 2D mask object predictions and raw LiDAR point clouds into full 3D bounding boxes of objects. Because the LiDAR point clouds are par…
PASS: An ImageNet replacement for self-supervised pretraining without humans
Yuki M. Asano, Christian Rupprecht, Andrew Zisserman +1
Computer vision has long relied on ImageNet and other large datasets of images sampled from the Internet for pretraining models. However, these datasets have ethical and technical…
Real Time Monocular Vehicle Velocity Estimation using Synthetic Data
Robert McCraith, Lukas Neumann, Andrea Vedaldi
Vision is one of the primary sensing modalities in autonomous driving. In this paper we look at the problem of estimating the velocity of road vehicles from a camera mounted on a m…
DensePose 3D: Lifting Canonical Surface Maps of Articulated Objects to the Third Dimension
Roman Shapovalov, David Novotny, Benjamin Graham +2
We tackle the problem of monocular 3D reconstruction of articulated objects like humans and animals. We contribute DensePose 3D, a method that can learn such reconstructions in a w…