1 paper
Yizhen Chen, Jie Wang, Lijian Lin +3
Vision-language alignment learning for video-text retrieval arouses a lot of attention in recent years. Most of the existing methods either transfer the knowledge of image-text pre…