1 citations · 1 across the 1 of their papers we have counts for
2 papers
cs.MM2024★ 1 cited
Open-Vocabulary Audio-Visual Semantic Segmentation
Ruohao Guo, Liao Qu, Dantong Niu +5
Audio-visual semantic segmentation (AVSS) aims to segment and classify sounding objects in videos with acoustic cues. However, most approaches operate on the close-set assumption a…
cs.CV2023
Audio-Visual Instance Segmentation
Ruohao Guo, Xianghua Ying, Yaru Chen +11
In this paper, we propose a new multi-modal task, termed audio-visual instance segmentation (AVIS), which aims to simultaneously identify, segment and track individual sounding obj…