1 paper
Zhangjing Yang, Dun Liu, Wensheng Cheng +2
Labeling pixel-wise object masks in videos is a resource-intensive and laborious process. Box-supervised Video Instance Segmentation (VIS) methods have emerged as a viable solution…