4 papers
Leverage Cross-Attention for End-to-End Open-Vocabulary Panoptic Reconstruction
Xuan Yu, Yuxuan Xie, Yili Liu +4
Open-vocabulary panoptic reconstruction offers comprehensive scene understanding, enabling advances in embodied robotics and photorealistic simulation. In this paper, we propose Pa…
Scale Disparity of Instances in Interactive Point Cloud Segmentation
Chenrui Han, Xuan Yu, Yuxuan Xie +5
Interactive point cloud segmentation has become a pivotal task for understanding 3D scenes, enabling users to guide segmentation models with simple interactions such as clicks, the…
Let Occ Flow: Self-Supervised 3D Occupancy Flow Prediction
Yili Liu, Linzhan Mou, Xuan Yu +4
Accurate perception of the dynamic environment is a fundamental task for autonomous driving and robot systems. This paper introduces Let Occ Flow, the first self-supervised work fo…
PanopticRecon: Leverage Open-vocabulary Instance Segmentation for Zero-shot Panoptic Reconstruction
Xuan Yu, Yili Liu, Chenrui Han +5
Panoptic reconstruction is a challenging task in 3D scene understanding. However, most existing methods heavily rely on pre-trained semantic segmentation models and known 3D object…