3 citations · 3 across the 3 of their papers we have counts for
3 papers
cs.MM2026
SAMOT: State-Aware Step Modulation and Optimal Transport Matching for Audio-Visual Instance Segmentation
Kai Peng, Yunzhe Shen, Miao Zhang +5
Audio-Visual Instance Segmentation (AVIS) aims to simultaneously classify, segment, and track sounding objects within video sequences. Unlike Audio-Visual Semantic Segmentation (AV…
cs.MM2026
Hear to See: Discerning Stateful Listening for Audio-Visual Instance Segmentation
Leiye Liu, Miao Zhang, Jiahong Jiang +7
Audio-visual instance segmentation (AVIS) requires accurately identifying and tracking individual sounding objects with pixel-level masks. Existing methods struggle to match overla…
cs.CV2021★ 3 cited
To be Critical: Self-Calibrated Weakly Supervised Learning for Salient Object Detection
Yongri Piao, Jian Wang, Miao Zhang +2
Weakly-supervised salient object detection (WSOD) aims to develop saliency models using image-level annotations. Despite of the success of previous works, explorations on an effect…