1 paper
Xin Wang, Jiawei Wu, Da Zhang +2
Although promising results have been achieved in video captioning, existing models are limited to the fixed inventory of activities in the training corpus, and do not generalize to…