3 papers
cs.CV2025
EPSegFZ: Efficient Point Cloud Semantic Segmentation for Few- and Zero-Shot Scenarios with Language Guidance
Jiahui Wang, Haiyue Zhu, Haoren Guo +3
Recent approaches for few-shot 3D point cloud semantic segmentation typically require a two-stage learning process, i.e., a pre-training stage followed by a few-shot training stage…
cs.CV2025
SingRef6D: Monocular Novel Object Pose Estimation with a Single RGB Reference
Jiahui Wang, Haiyue Zhu, Haoren Guo +3
Recent 6D pose estimation methods demonstrate notable performance but still face some practical limitations. For instance, many of them rely heavily on sensor depth, which may fail…
cs.RO2025
VLM-UDMC: VLM-Enhanced Unified Decision-Making and Motion Control for Urban Autonomous Driving
Haichao Liu, Haoren Guo, Pei Liu +4
Scene understanding and risk-aware attentions are crucial for human drivers to make safe and effective driving decisions. To imitate this cognitive ability in urban autonomous driv…