11 citations · 15 across the 2 of their papers we have counts for
3 papers
cs.CV2022★ 11 cited
Cross-view Transformers for real-time Map-view Semantic Segmentation
Brady Zhou, Philipp Krähenbühl
We present cross-view transformers, an efficient attention-based model for map-view semantic segmentation from multiple cameras. Our architecture implicitly learns a mapping from i…
cs.RO2019★ 4 cited
Learning by Cheating
Dian Chen, Brady Zhou, Vladlen Koltun +1
Vision-based urban driving is hard. The autonomous system needs to learn to perceive the world and act in it. We show that this challenging learning problem can be simplified by de…
cs.CV2019
Does computer vision matter for action?
Brady Zhou, Philipp Krähenbühl, Vladlen Koltun
Computer vision produces representations of scene content. Much computer vision research is predicated on the assumption that these intermediate representations are useful for acti…