4 papers
Progressive Prompt-Guided Cross-Modal Reasoning for Referring Image Segmentation
Jiachen Li, Hongyun Wang, Jinyu Xu +5
Referring image segmentation aims to localize and segment a target object in an image based on a free-form referring expression. The core challenge lies in effectively bridging lin…
RWKV-PCSSC: Exploring RWKV Model for Point Cloud Semantic Scene Completion
Wenzhe He, Xiaojun Chen, Wentang Chen +3
Semantic Scene Completion (SSC) aims to generate a complete semantic scene from an incomplete input. Existing approaches often employ dense network architectures with a high parame…
GauSSmart: Enhanced 3D Reconstruction through 2D Foundation Models and Geometric Filtering
Alexander Valverde, Brian Xu, Yuyin Zhou +2
Scene reconstruction has emerged as a central challenge in computer vision, with approaches such as Neural Radiance Fields (NeRF) and Gaussian Splatting achieving remarkable progre…
LGD: Leveraging Generative Descriptions for Zero-Shot Referring Image Segmentation
Jiachen Li, Qing Xie, Renshu Gu +3
Zero-shot referring image segmentation aims to locate and segment the target region based on a referring expression, with the primary challenge of aligning and matching semantics a…