3 papers
cs.CV2026
Seek to Segment: Active Perception for Panoramic Referring Segmentation
Song Tang, Shuming Hu, Xincheng Shuai +2
Existing referring segmentation models passively process static images captured from fixed perspectives, limiting their applicability in Embodied AI, where agents must perform acti…
cs.CV2026
ROSE: Retrieval-Oriented Segmentation Enhancement
Song Tang, Guangquan Jie, Henghui Ding +1
Existing segmentation models based on multimodal large language models (MLLMs), such as LISA, often struggle with novel or emerging entities due to their inability to incorporate u…
cs.CV2025
Multimodal Referring Segmentation: A Survey
Henghui Ding, Song Tang, Shuting He +3
Multimodal referring segmentation aims to segment target objects in visual scenes, such as images, videos, and 3D scenes, based on referring expressions in text or audio format. Th…