4 papers
Leverage Cross-Attention for End-to-End Open-Vocabulary Panoptic Reconstruction
Xuan Yu, Yuxuan Xie, Yili Liu +4
Open-vocabulary panoptic reconstruction offers comprehensive scene understanding, enabling advances in embodied robotics and photorealistic simulation. In this paper, we propose Pa…
Let Occ Flow: Self-Supervised 3D Occupancy Flow Prediction
Yili Liu, Linzhan Mou, Xuan Yu +4
Accurate perception of the dynamic environment is a fundamental task for autonomous driving and robot systems. This paper introduces Let Occ Flow, the first self-supervised work fo…
Scale Disparity of Instances in Interactive Point Cloud Segmentation
Chenrui Han, Xuan Yu, Yuxuan Xie +5
Interactive point cloud segmentation has become a pivotal task for understanding 3D scenes, enabling users to guide segmentation models with simple interactions such as clicks, the…
PanopticRecon: Leverage Open-vocabulary Instance Segmentation for Zero-shot Panoptic Reconstruction
Xuan Yu, Yili Liu, Chenrui Han +5
Panoptic reconstruction is a challenging task in 3D scene understanding. However, most existing methods heavily rely on pre-trained semantic segmentation models and known 3D object…