1 citations · 1 across the 1 of their papers we have counts for
1 paper · 1 filter
Ruohao Guo, Xianghua Ying, Yaru Chen +11
In this paper, we propose a new multi-modal task, termed audio-visual instance segmentation (AVIS), which aims to simultaneously identify, segment and track individual sounding obj…