1 paper
Minghao Zhu, Xiao Lin, Mengxian Hu +5
Adapting CLIP for open-vocabulary video recognition necessitates a delicate balance between newly acquired video knowledge and the pretrained generalization. While existing studies…