4 papers
FPC-VLA: A Vision-Language-Action Framework with a Supervisor for Failure Prediction and Correction
Yifan Yang, Zhixiang Duan, Tianshi Xie +8
Robotic manipulation is a fundamental component of automation. However, traditional perception-planning pipelines often fall short in open-ended tasks due to limited flexibility, w…
MK-Pose: Category-Level Object Pose Estimation via Multimodal-Based Keypoint Learning
Yifan Yang, Peili Song, Enfan Lan +2
Category-level object pose estimation, which predicts the pose of objects within a known category without prior knowledge of individual instances, is essential in applications like…
Simulating Automotive Radar with Lidar and Camera Inputs
Peili Song, Dezhen Song, Yifan Yang +2
Low-cost millimeter automotive radar has received more and more attention due to its ability to handle adverse weather and lighting conditions in autonomous driving. However, the l…
PS6D: Point Cloud Based Symmetry-Aware 6D Object Pose Estimation in Robot Bin-Picking
Yifan Yang, Zhihao Cui, Qianyi Zhang +1
6D object pose estimation holds essential roles in various fields, particularly in the grasping of industrial workpieces. Given challenges like rust, high reflectivity, and absent…