3k citations · 3.4k across the 8 of their papers we have counts for
18 papers
Global Tracking Transformers
Xingyi Zhou, Tianwei Yin, Vladlen Koltun +1
We present a novel transformer-based architecture for global multi-object tracking. Our network takes a short sequence of frames as input and produces global trajectories for all o…
Towards Long-Form Video Understanding
Chao-Yuan Wu, Philipp Krähenbühl
Our world offers a never-ending stream of visual stimuli, yet today's vision systems only accurately recognize patterns within a few seconds. These systems understand the present,…
Learning to drive from a world on rails
Dian Chen, Vladlen Koltun, Philipp Krähenbühl
We learn an interactive vision-based driving policy from pre-recorded driving logs via a model-based approach. A forward model of the world supervises a driving policy that predict…
Probabilistic two-stage detection
Xingyi Zhou, Vladlen Koltun, Philipp Krähenbühl
We develop a probabilistic interpretation of two-stage object detection. We show that this probabilistic interpretation motivates a number of common empirical training practices. I…
Lossless Image Compression through Super-Resolution
Sheng Cao, Chao-Yuan Wu, Philipp Krähenbühl
We introduce a simple and efficient lossless image compression algorithm. We store a low resolution version of an image as raw pixels, followed by several iterations of lossless su…
Tracking Objects as Points
Xingyi Zhou, Vladlen Koltun, Philipp Krähenbühl
Tracking has traditionally been the art of following interest points through space and time. This changed with the rise of powerful deep networks. Nowadays, tracking is dominated b…