4 papers
Agreement-Based Audio-Visual Segmentation:Champion Report for the MeViS-Audio Track in the 8th LSVOS Challenge
Yiwen Ren, Jianing Liu, Yingxin Wang +4
The MeViS-Audio track asks a system to segment the objects described by a spoken motion expression throughout a video and to return empty masks when the described target is absent.…
Turbulence-Robust Dynamic Object Segmentation with Multi-Signal Priors and SAM2 Refinement
Bolian Peng, Ying Tang, Xu Liu +2
This technical report presents our solution for the CVPR 2026 UG2+ Challenge Track 3: Dynamic Object Segmentation in Turbulence (DOST). We design a training-free multi-signal segme…
PVUW 2025 Challenge Report: Advances in Pixel-level Understanding of Complex Videos in the Wild
Henghui Ding, Chang Liu, Nikhila Ravi +33
This report provides a comprehensive overview of the 4th Pixel-level Video Understanding in the Wild (PVUW) Challenge, held in conjunction with CVPR 2025. It summarizes the challen…
FVOS for MOSE Track of 4th PVUW Challenge: 3rd Place Solution
Mengjiao Wang, Junpei Zhang, Xu Liu +2
Video Object Segmentation (VOS) is one of the most fundamental and challenging tasks in computer vision and has a wide range of applications. Most existing methods rely on spatiote…