1.6k citations · 2.3k across the 7 of their papers we have counts for
6 papers · 1 filter
Everybody Tracking Every Body
Daeyun Shin, Yunhan Zhao, Shu Kong +2
We address the problem of 3D body pose estimation of multiple interacting people from their egocentric views with centralized coordination. Each individual wears a camera recording…
Joint Depth Prediction and Semantic Segmentation with Multi-View SAM
Mykhailo Shvets, Dongxu Zhao, Marc Niethammer +2
Multi-task approaches to joint depth and segmentation prediction are well-studied for monocular images. Yet, predictions from a single-view are inherently limited, while multiple v…
Segment Anything
Alexander Kirillov, Eric Mintun, Nikhila Ravi +9
We introduce the Segment Anything (SA) project: a new task, model, and dataset for image segmentation. Using our efficient model in a data collection loop, we built the largest seg…
DSSD : Deconvolutional Single Shot Detector
Cheng-Yang Fu, Wei Liu, Ananth Ranga +2
The main contribution of this paper is an approach for introducing additional context into state-of-the-art general object detection. To achieve this we first combine a state-of-th…
Modeling Context in Referring Expressions
Licheng Yu, Patrick Poirson, Shan Yang +2
Humans refer to objects in their environments all the time, especially in dialogue with other people. We explore generating and comprehending natural language referring expressions…
ImageNet Large Scale Visual Recognition Challenge
Olga Russakovsky, Jia Deng, Hao Su +9
The ImageNet Large Scale Visual Recognition Challenge is a benchmark in object category classification and detection on hundreds of object categories and millions of images. The ch…