52 citations · 113 across the 26 of their papers we have counts for
1 paper · 2 filters
Ziqi Zhang, Yaya Shi, Chunfeng Yuan +4
Taking full advantage of the information from both vision and language is critical for the video captioning task. Existing models lack adequate visual representation due to the neg…