1 citations · 2 across the 5 of their papers we have counts for
1 paper · 2 filters
Jinze Lv, Jian Chen, Zi Long +2
Most existing multimodal machine translation (MMT) datasets are predominantly composed of static images or short video clips, lacking extensive video data across diverse domains an…