6 citations · 6 across the 2 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2020★ 6 cited
Multi-modal Feature Fusion with Feature Attention for VATEX Captioning Challenge 2020
Ke Lin, Zhuoxin Gan, Liwei Wang
This report describes our model for VATEX Captioning Challenge 2020. First, to gather information from multiple domains, we extract motion, appearance, semantic and audio features.…
cs.CV2019
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling
Haoran Chen, Ke Lin, Alexander Maye +2
Given the features of a video, recurrent neural networks can be used to automatically generate a caption for the video. Existing methods for video captioning have at least three li…