1 paper
Sungho Moon, Seunghun Lee, Jiwan Seo +1
We propose Context-aware Video-text Alignment (CVA), a novel framework to address a significant challenge in video temporal grounding: achieving temporally sensitive video-text ali…