70 citations · 206 across the 16 of their papers we have counts for
1 paper · 2 filters
Jie Lei, Liwei Wang, Yelong Shen +3
Generating multi-sentence descriptions for videos is one of the most challenging captioning tasks due to its high requirements for not only visual relevance but also discourse-base…