747 citations · 2.1k across the 32 of their papers we have counts for
7 papers · 1 filter
Towards Robust Image Classification Using Sequential Attention Models
Daniel Zoran, Mike Chrzanowski, Po-Sen Huang +3
In this paper we propose to augment a modern neural-network architecture with an attention model inspired by human perception. Specifically, we adversarially train and analyze a ne…
CLEVRER: CoLlision Events for Video REpresentation and Reasoning
Kexin Yi, Chuang Gan, Yunzhu Li +4
The ability to reason about temporal and causal events from videos lies at the core of human intelligence. Most video reasoning benchmarks, however, focus on pattern recognition fr…
A Hierarchical Probabilistic U-Net for Modeling Multi-Scale Ambiguities
Simon A. A. Kohl, Bernardino Romera-Paredes, Klaus H. Maier-Hein +5
Medical imaging only indirectly measures the molecular identity of the tissue within each voxel, which often produces only ambiguous image evidence for target measures of interest,…
The Neuro-Symbolic Concept Learner: Interpreting Scenes, Words, and Sentences From Natural Supervision
Jiayuan Mao, Chuang Gan, Pushmeet Kohli +2
We propose the Neuro-Symbolic Concept Learner (NS-CL), a model that learns visual concepts, words, and semantic parsing of sentences without explicit supervision on any of them; in…
Efficient Relaxations for Dense CRFs with Sparse Higher Order Potentials
Thomas Joy, Alban Desmaison, Thalaiyasingam Ajanthan +5
Dense conditional random fields (CRFs) have become a popular framework for modelling several problems in computer vision such as stereo correspondence and multi-class semantic segm…
Learning to Navigate the Energy Landscape
Julien Valentin, Angela Dai, Matthias Nießner +4
In this paper, we present a novel and efficient architecture for addressing computer vision problems that use `Analysis by Synthesis'. Analysis by synthesis involves the minimizati…