1 citations · 1 across the 1 of their papers we have counts for
1 paper
Zhiyu Tan, Xiaomeng Yang, Luozheng Qin +1
The quality of video-text pairs fundamentally determines the upper bound of text-to-video models. Currently, the datasets used for training these models suffer from significant sho…