1 citations · 1 across the 3 of their papers we have counts for
1 paper · 1 filter
Dongjie Yang, Suyuan Huang, Chengqiang Lu +5
Advancements in multimodal learning, particularly in video understanding and generation, require high-quality video-text datasets for improved model performance. Vript addresses th…