1 paper
Rui Lin, Chuanming Wang, Huadong Ma
With the rapid development of pre-training technologies, adapting large-scale Vision-Language Models (VLMs) for video understanding \emph{\ie} image-to-video transfer learning has…