4 papers · 1 filter
3DWay: Generalizing Robot Manipulation via 3D Consistent Waypoints
Ziqin Huang, Yingyue Li, Chenyangguang Zhang +6
Intermediate representations are key to bridging the modality gap between generalizable manipulation policies and large-scale pretrained vision-language models (VLMs). Among these,…
Functional-SLAM: Interaction-Aware Mapping with Online Functional Scene Graphs
Xinggang Hu, Chenyangguang Zhang, Zihan Zhu +3
Existing SLAM systems lack modeling of the functional relations required for fine-grained robotic interaction. Functional 3D scene graphs can represent relations between objects an…
Hierarchical and Holistic Open-Vocabulary Functional 3D Scene Graphs for Indoor Spaces
Xinggang Hu, Chenyangguang Zhang, Alexandros Delitzas +4
Functional 3D scene graphs offer a versatile and flexible representation for 3D scene understanding and robotic manipulation, defined by object nodes, interactive elements, and fun…
LanPose: Language-Instructed 6D Object Pose Estimation for Robotic Assembly
Bowen Fu, Sek Kun Leong, Yan Di +2
Comprehending natural language instructions is a critical skill for robots to cooperate effectively with humans. In this paper, we aim to learn 6D poses for roboticassembly by natu…