1 paper
Seung hee Choi, MinJu Jeon, Hyunwoo Oh +2
Existing retrieval-augmented approaches for Dense Video Captioning (DVC) often fail to achieve accurate temporal segmentation aligned with true event boundaries, as they rely on he…