24 citations · 32 across the 6 of their papers we have counts for
7 papers
Hierarchical Multimodal Transformer to Summarize Videos
Bin Zhao, Maoguo Gong, Xuelong Li
Although video summarization has achieved tremendous success benefiting from Recurrent Neural Networks (RNN), RNN-based methods neglect the global dependencies and multi-hop relati…
EA-Net: Edge-Aware Network for Flow-based Video Frame Interpolation
Bin Zhao, Xuelong Li
Video frame interpolation can up-convert the frame rate and enhance the video quality. In recent years, although the interpolation performance has achieved great success, image blu…
AudioVisual Video Summarization
Bin Zhao, Maoguo Gong, Xuelong Li
Audio and vision are two main modalities in video data. Multimodal learning, especially for audiovisual learning, has drawn considerable attention recently, which can boost the per…
Reconstructive Sequence-Graph Network for Video Summarization
Bin Zhao, Haopeng Li, Xiaoqiang Lu +1
Exploiting the inner-shot and inter-shot dependencies is essential for key-shot based video summarization. Current approaches mainly devote to modeling the video as a frame sequenc…
Weather GAN: Multi-Domain Weather Translation Using Generative Adversarial Networks
Xuelong Li, Kai Kou, Bin Zhao
In this paper, a new task is proposed, namely, weather translation, which refers to transferring weather conditions of the image from one category to another. It is important for p…
Generating Masks from Boxes by Mining Spatio-Temporal Consistencies in Videos
Bin Zhao, Goutam Bhat, Martin Danelljan +2
Segmenting objects in videos is a fundamental computer vision task. The current deep learning based paradigm offers a powerful, but data-hungry solution. However, current datasets…