1 paper
Minghan Li, Shuai Li, Wangmeng Xiang +1
While impressive progress has been achieved, video instance segmentation (VIS) methods with per-clip input often fail on challenging videos with occluded objects and crowded scenes…