70 citations · 187 across the 15 of their papers we have counts for
1 paper · 1 filter
Jie Lei, Liwei Wang, Yelong Shen +3
Generating multi-sentence descriptions for videos is one of the most challenging captioning tasks due to its high requirements for not only visual relevance but also discourse-base…