3 papers
cs.CV2026
Seeing Realism from Simulation: Efficient Video Transfer for Vision-Language-Action Data Augmentation
Chenyu Hui, Xiaodi Huang, Siyu Xu +5
Vision-language-action (VLA) models typically rely on large-scale real-world videos, whereas simulated data, despite being inexpensive and highly parallelizable to collect, often s…
cs.CV2025
Shape Distribution Matters: Shape-specific Mixture-of-Experts for Amodal Segmentation under Diverse Occlusions
Zhixuan Li, Yujia Liu, Chen Hui +3
Amodal segmentation targets to predict complete object masks, covering both visible and occluded regions. This task poses significant challenges due to complex occlusions and extre…
cs.CV2025
Single Point, Full Mask: Velocity-Guided Level Set Evolution for End-to-End Amodal Segmentation
Zhixuan Li, Yujia Liu, Chen Hui +1
Amodal segmentation aims to recover complete object shapes, including occluded regions with no visual appearance, whereas conventional segmentation focuses solely on visible areas.…