1 paper
Mingkai Tian, Guorong Li, Yuankai Qi +4
Zero-shot video captioning requires that a model generate high-quality captions without human-annotated video-text pairs for training. State-of-the-art approaches to the problem le…