79 citations · 342 across the 26 of their papers we have counts for
1 paper · 2 filters
David M. Chan, Sudheendra Vijayanarasimhan, David A. Ross +1
Automatic video captioning aims to train models to generate text descriptions for all segments in a video, however, the most effective approaches require large amounts of manual an…