650 citations · 762 across the 17 of their papers we have counts for
10 papers · 1 filter
Diverse Image Captioning with Context-Object Split Latent Spaces
Shweta Mahajan, Stefan Roth
Diverse image captioning models aim to learn one-to-many mappings that are innate to cross-domain datasets, such as of images and texts. Current methods for this task are based on…
MOTChallenge: A Benchmark for Single-Camera Multiple Target Tracking
Patrick Dendorfer, Aljoša Ošep, Anton Milan +5
Standardized benchmarks have been crucial in pushing the performance of computer vision algorithms, especially since the advent of deep learning. Although leaderboards should not b…
LR-CNN: Local-aware Region CNN for Vehicle Detection in Aerial Imagery
Wentong Liao, Xiang Chen, Jingfeng Yang +4
State-of-the-art object detection approaches such as Fast/Faster R-CNN, SSD, or YOLO have difficulties detecting dense, small targets with arbitrary orientation in large aerial ima…
Single-Stage Semantic Segmentation from Image Labels
Nikita Araslanov, Stefan Roth
Recent years have seen a rapid growth in new approaches improving the accuracy of semantic segmentation in a weakly supervised setting, i.e. with only image-level labels available…
Self-Supervised Monocular Scene Flow Estimation
Junhwa Hur, Stefan Roth
Scene flow estimation has been receiving increasing attention for 3D environment perception. Monocular scene flow estimation -- obtaining 3D structure and 3D motion from two tempor…
Normalizing Flows with Multi-Scale Autoregressive Priors
Shweta Mahajan, Apratim Bhattacharyya, Mario Fritz +2
Flow-based generative models are an important class of exact inference models that admit efficient inference and sampling for image synthesis. Owing to the efficiency constraints o…