activity
20152023
most citedBLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models

924 citations · 2.3k across the 66 of their papers we have counts for

collaborators
Showing 2017 · cs.CVShow all

14 papers · 2 filters

cs.CV2017★ 12 cited

Im2Pano3D: Extrapolating 360 Structure and Semantics Beyond the Field of View

Shuran Song, Andy Zeng, Angel X. Chang +3

We present Im2Pano3D, a convolutional neural network that generates a dense prediction of 3D structure and a probability distribution of semantic labels for a full 360 panoramic vi…

cs.CV2017

CAR-Net: Clairvoyant Attentive Recurrent Network

Amir Sadeghian, Ferdinand Legros, Maxime Voisin +3

We present an interpretable framework for path prediction that leverages dependencies between agents' behaviors and their spatial navigation environment. We exploit two sources of…

cs.CV2017

Adversarial Feature Augmentation for Unsupervised Domain Adaptation

Riccardo Volpi, Pietro Morerio, Silvio Savarese +1

Recent works showed that Generative Adversarial Networks (GANs) can be successfully applied in unsupervised domain adaptation, where, given a labeled source dataset and an unlabele…

cs.CV2017

Recurrent Autoregressive Networks for Online Multi-Object Tracking

Kuan Fang, Yu Xiang, Xiaocheng Li +1

The main challenge of online multi-object tracking is to reliably associate object trajectories with detections in each video frame based on their tracking history. In this work, w…

cs.CV2017★ 53 cited

Large-Scale 3D Shape Reconstruction and Segmentation from ShapeNet Core55

Li Yi, Lin Shao, Manolis Savva +47

We introduce a large-scale 3D shape understanding benchmark using data and annotation from ShapeNet 3D object database. The benchmark consists of two tasks: part-level segmentation…

cs.CV2017★ 75 cited

Generic 3D Representation via Pose Estimation and Matching

Amir R. Zamir, Tilman Wekel, Pulkit Argrawal +3

Though a large body of computer vision research has investigated developing generic semantic representations, efforts towards developing a similar representation for 3D has been li…