1 paper
Wenqi Zhu, Jiale Cao, Jin Xie +2
Open-vocabulary video instance segmentation strives to segment and track instances belonging to an open set of categories in a videos. The vision-language model Contrastive Languag…