1 paper · 1 filter
Jian Chen, JinZe Lv, Zi Long +1
Video-guided Multimodal Translation (VMT) has advanced significantly in recent years. However, most existing methods rely on locally aligned video segments paired one-to-one with s…