4 citations · 4 across the 2 of their papers we have counts for
2 papers
cs.CV2021★ 4 cited
CLIP4Caption ++: Multi-CLIP for Video Caption
Mingkang Tang, Zhanyu Wang, Zhaoyang Zeng +2
This report describes our solution to the VALUE Challenge 2021 in the captioning task. Our solution, named CLIP4Caption++, is built on X-Linear/X-Transformer, which is an advanced…
cs.CV2021
CLIP4Caption: CLIP for Video Caption
Mingkang Tang, Zhanyu Wang, Zhenhua Liu +3
Video captioning is a challenging task since it requires generating sentences describing various diverse and complex videos. Existing video captioning models lack adequate visual r…