activity
20192023
most citedLocalizing Unseen Activities in Video via Image Query

7 citations · 10 across the 4 of their papers we have counts for

collaborators

5 papers

cs.MM2023

Weakly-Supervised Video Moment Retrieval via Regularized Two-Branch Proposal Networks with Erasing Mechanism

Haoyuan Li, Zhou Zhao, Zhu Zhang +1

Video moment retrieval is to identify the target moment according to the given sentence in an untrimmed video. Due to temporal boundary annotations of the video are extremely time-…

cs.CV20212 cited

SimulLR: Simultaneous Lip Reading Transducer with Attention-Guided Adaptive Memory

Zhijie Lin, Zhou Zhao, Haoyuan Li +4

Lip reading, aiming to recognize spoken sentences according to the given video of lip movements without relying on the audio stream, has attracted great interest due to its applica…

cs.IR2019

Cross-Modal Interaction Networks for Query-Based Moment Retrieval in Videos

Zhu Zhang, Zhijie Lin, Zhou Zhao +1

Query-based moment retrieval aims to localize the most relevant moment in an untrimmed video according to the given natural language query. Existing works often only focus on one a…

cs.CV20197 cited

Localizing Unseen Activities in Video via Image Query

Zhu Zhang, Zhou Zhao, Zhijie Lin +2

Action localization in untrimmed videos is an important topic in the field of video understanding. However, existing action localization methods are restricted to a pre-defined set…

cs.CV20191 cited

Open-Ended Long-Form Video Question Answering via Hierarchical Convolutional Self-Attention Networks

Zhu Zhang, Zhou Zhao, Zhijie Lin +2

Open-ended video question answering aims to automatically generate the natural-language answer from referenced video contents according to the given question. Currently, most exist…