1 paper
Baoyao Yang, Wanyun Li, Dixin Chen +3
This paper introduces VideoMind, a video-centric omni-modal dataset designed for deep video content cognition and enhanced multi-modal feature representation. The dataset comprises…