1 paper
Arian Kheirandish, Fardin Ayar, Ehsan Javanmardi +2
Video instance segmentation (VIS) requires models to detect, segment, and track object identities across frames, and most methods enforce temporal consistency through video-level s…