32 citations · 32 across the 5 of their papers we have counts for
1 paper · 1 filter
Xin Wang, Jiawei Wu, Junkun Chen +3
We present a new large-scale multilingual video description dataset, VATEX, which contains over 41,250 videos and 825,000 captions in both English and Chinese. Among the captions,…