94 citations · 391 across the 20 of their papers we have counts for
4 papers · 1 filter
Learning to Compose Visual Relations
Nan Liu, Shuang Li, Yilun Du +2
The visual world around us can be described as a structured set of objects and their associated relations. An image of a room may be conjured given only the description of the unde…
Unsupervised Learning of Compositional Energy Concepts
Yilun Du, Shuang Li, Yash Sharma +2
Humans are able to rapidly understand scenes by utilizing concepts extracted from prior experience. Such concepts are diverse, and include global scene descriptors, such as the wea…
Weakly Supervised Human-Object Interaction Detection in Video via Contrastive Spatiotemporal Regions
Shuang Li, Yilun Du, Antonio Torralba +2
We introduce the task of weakly supervised learning for detecting human and object interactions in videos. Our task poses unique challenges as a system does not know what types of…
3D Neural Scene Representations for Visuomotor Control
Yunzhu Li, Shuang Li, Vincent Sitzmann +2
Humans have a strong intuitive understanding of the 3D environment around us. The mental model of the physics in our brain applies to objects of different materials and enables us…