1 paper
Jisheng Dang, Yizhou Zhang, Hao Ye +6
Fine-grained video captioning aims to generate detailed, temporally coherent descriptions of video content. However, existing methods struggle to capture subtle video dynamics and…