23 citations · 42 across the 3 of their papers we have counts for
3 papers
cs.CV2019
TruNet: Short Videos Generation from Long Videos via Story-Preserving Truncation
Fan Yang, Xiao Liu, Dongliang He +5
In this work, we introduce a new problem, named as {\em story-preserving long video truncation}, that requires an algorithm to automatically truncate a long-duration video into mul…
cs.CV2019★ 19 cited
Read, Watch, and Move: Reinforcement Learning for Temporally Grounding Natural Language Descriptions in Videos
Dongliang He, Xiang Zhao, Jizhou Huang +3
The task of video grounding, which temporally localizes a natural language description in a video, plays an important role in understanding videos. Existing studies have adopted st…
cs.CV2018★ 23 cited
StNet: Local and Global Spatial-Temporal Modeling for Action Recognition
Dongliang He, Zhichao Zhou, Chuang Gan +5
Despite the success of deep learning for static image understanding, it remains unclear what are the most effective network architectures for the spatial-temporal modeling in video…