7 citations · 10 across the 4 of their papers we have counts for
5 papers
Weakly-Supervised Video Moment Retrieval via Regularized Two-Branch Proposal Networks with Erasing Mechanism
Haoyuan Li, Zhou Zhao, Zhu Zhang +1
Video moment retrieval is to identify the target moment according to the given sentence in an untrimmed video. Due to temporal boundary annotations of the video are extremely time-…
SimulLR: Simultaneous Lip Reading Transducer with Attention-Guided Adaptive Memory
Zhijie Lin, Zhou Zhao, Haoyuan Li +4
Lip reading, aiming to recognize spoken sentences according to the given video of lip movements without relying on the audio stream, has attracted great interest due to its applica…
Cross-Modal Interaction Networks for Query-Based Moment Retrieval in Videos
Zhu Zhang, Zhijie Lin, Zhou Zhao +1
Query-based moment retrieval aims to localize the most relevant moment in an untrimmed video according to the given natural language query. Existing works often only focus on one a…
Localizing Unseen Activities in Video via Image Query
Zhu Zhang, Zhou Zhao, Zhijie Lin +2
Action localization in untrimmed videos is an important topic in the field of video understanding. However, existing action localization methods are restricted to a pre-defined set…
Open-Ended Long-Form Video Question Answering via Hierarchical Convolutional Self-Attention Networks
Zhu Zhang, Zhou Zhao, Zhijie Lin +2
Open-ended video question answering aims to automatically generate the natural-language answer from referenced video contents according to the given question. Currently, most exist…