26 citations · 31 across the 5 of their papers we have counts for
6 papers · 1 filter
Brain-Inspired Inference on Missing Video Sequence
Weimian Li, Baoyang Chen, Wenmin Wang
In this paper, we propose a novel end-to-end architecture that could generate a variety of plausible video sequences correlating two given discontinuous frames. Our work is inspire…
Adaptively Aligned Image Captioning via Adaptive Attention Time
Lun Huang, Wenmin Wang, Yaxian Xia +1
Recent neural models for image captioning usually employ an encoder-decoder framework with an attention mechanism. However, the attention mechanism in such a framework aligns one s…
Attention on Attention for Image Captioning
Lun Huang, Wenmin Wang, Jie Chen +1
Attention mechanisms are widely used in current encoder/decoder frameworks of image captioning, where a weighted average on encoded vectors is generated at each time step to guide…
ParNet: Position-aware Aggregated Relation Network for Image-Text matching
Yaxian Xia, Lun Huang, Wenmin Wang +1
Exploring fine-grained relationship between entities(e.g. objects in image or words in sentence) has great contribution to understand multimedia content precisely. Previous attenti…
Video Imagination from a Single Image with Transformation Generation
Baoyang Chen, Wenmin Wang, Jinzhuo Wang +1
In this work, we focus on a challenging task: synthesizing multiple imaginary videos given a single image. Major problems come from high dimensionality of pixel space and the ambig…
Long-Term Video Interpolation with Bidirectional Predictive Network
Xiongtao Chen, Wenmin Wang, Jinzhuo Wang +2
This paper considers the challenging task of long-term video interpolation. Unlike most existing methods that only generate few intermediate frames between existing adjacent ones,…