5 citations · 7 across the 2 of their papers we have counts for
4 papers
Controllable Video Captioning with POS Sequence Guidance Based on Gated Fusion Network
Bairui Wang, Lin Ma, Wei Zhang +3
In this paper, we propose to guide the video caption generation with Part-of-Speech (POS) information, based on a gated fusion of multiple representations of input videos. We const…
Reconstruct and Represent Video Contents for Captioning via Reinforcement Learning
Wei Zhang, Bairui Wang, Lin Ma +1
In this paper, the problem of describing visual contents of a video sequence with natural language is addressed. Unlike previous video captioning work mainly exploiting the cues of…
Hierarchical Photo-Scene Encoder for Album Storytelling
Bairui Wang, Lin Ma, Wei Zhang +2
In this paper, we propose a novel model with a hierarchical photo-scene encoder and a reconstructor for the task of album storytelling. The photo-scene encoder contains two sub-enc…
Reconstruction Network for Video Captioning
Bairui Wang, Lin Ma, Wei Zhang +1
In this paper, the problem of describing visual contents of a video sequence with natural language is addressed. Unlike previous video captioning work mainly exploiting the cues of…