2 papers
cs.CV2020
Delving Deeper into the Decoder for Video Captioning
Haoran Chen, Jianmin Li, Xiaolin Hu
Video captioning is an advanced multi-modal task which aims to describe a video clip using a natural language sentence. The encoder-decoder framework is the most popular paradigm f…
cs.CV2019
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling
Haoran Chen, Ke Lin, Alexander Maye +2
Given the features of a video, recurrent neural networks can be used to automatically generate a caption for the video. Existing methods for video captioning have at least three li…