80 citations · 317 across the 18 of their papers we have counts for
4 papers · 1 filter
IMP: Instance Mask Projection for High Accuracy Semantic Segmentation of Things
Cheng-Yang Fu, Tamara L. Berg, Alexander C. Berg
In this work, we present a new operator, called Instance Mask Projection (IMP), which projects a predicted Instance Segmentation as a new feature for semantic segmentation. It also…
Multi-Target Embodied Question Answering
Licheng Yu, Xinlei Chen, Georgia Gkioxari +3
Embodied Question Answering (EQA) is a relatively new task where an agent is asked to answer questions about its environment from egocentric perception. EQA makes the fundamental a…
TVQA+: Spatio-Temporal Grounding for Video Question Answering
Jie Lei, Licheng Yu, Tamara L. Berg +1
We present the task of Spatio-Temporal Video Question Answering, which requires intelligent systems to simultaneously retrieve relevant moments and detect referenced visual concept…
Dance Dance Generation: Motion Transfer for Internet Videos
Yipin Zhou, Zhaowen Wang, Chen Fang +2
This work presents computational methods for transferring body movements from one person to another with videos collected in the wild. Specifically, we train a personalized model o…