1 paper
Ruohao Guo, Liao Qu, Dantong Niu +5
Audio-visual semantic segmentation (AVSS) aims to segment and classify sounding objects in videos with acoustic cues. However, most approaches operate on the close-set assumption a…