1 paper
Kewei Tu, Meng Meng, Mun Wai Lee +2
We propose a framework for parsing video and text jointly for understanding events and answering user queries. Our framework produces a parse graph that represents the compositiona…