26 citations · 77 across the 4 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2021★ 26 cited
Discriminative Latent Semantic Graph for Video Captioning
Yang Bai, Junyan Wang, Yang Long +4
Video captioning aims to automatically generate natural language sentences that can describe the visual contents of a given video. Existing generative models like encoder-decoder f…
cs.CV2020★ 21 cited
Query Twice: Dual Mixture Attention Meta Learning for Video Summarization
Junyan Wang, Yang Bai, Yang Long +4
Video summarization aims to select representative frames to retain high-level information, which is usually solved by predicting the segment-wise importance score via a softmax fun…