2 papers
cs.CV2026
Video-guided Machine Translation with Global Video Context
Jian Chen, JinZe Lv, Zi Long +1
Video-guided Multimodal Translation (VMT) has advanced significantly in recent years. However, most existing methods rely on locally aligned video segments paired one-to-one with s…
cs.CL2025
TopicVD: A Topic-Based Dataset of Video-Guided Multimodal Machine Translation for Documentaries
Jinze Lv, Jian Chen, Zi Long +2
Most existing multimodal machine translation (MMT) datasets are predominantly composed of static images or short video clips, lacking extensive video data across diverse domains an…