1 paper
Ping Li, Tao Wang, Xinkui Zhao +2
Video captioning generate a sentence that describes the video content. Existing methods always require a number of captions (\eg, 10 or 20) per video to train the model, which is q…