13 citations · 13 across the 1 of their papers we have counts for
2 papers
cs.CV2015★ 13 cited
A Restricted Visual Turing Test for Deep Scene and Event Understanding
Hang Qi, Tianfu Wu, Mun-Wai Lee +1
This paper presents a restricted visual Turing test (VTT) for story-line based deep understanding in long-term and multi-camera captured videos. Given a set of videos of a scene (s…
cs.CV2013
Joint Video and Text Parsing for Understanding Events and Answering Queries
Kewei Tu, Meng Meng, Mun Wai Lee +2
We propose a framework for parsing video and text jointly for understanding events and answering user queries. Our framework produces a parse graph that represents the compositiona…