4 papers
Co-Fusion4D: Spatio-temporal Collaborative Fusion for Robust 3D Object Detection
Wenxuan Li, Qin Zou, Shoubing Chen +3
In autonomous driving, 3D object detection is essential for accurate perception and reliable decision-making. However, object motion and ego-motion often induce cross-frame spatiot…
CLASP: Closed-loop Asynchronous Spatial Perception for Open-vocabulary Desktop Object Grasping
Yiran Ling, Wenxuan Li, Siying Dong +5
Robot grasping of desktop object is widely used in intelligent manufacturing, logistics, and agriculture.Although vision-language models (VLMs) show strong potential for robotic ma…
PIN-WM: Learning Physics-INformed World Models for Non-Prehensile Manipulation
Wenxuan Li, Hang Zhao, Zhiyuan Yu +4
While non-prehensile manipulation (e.g., controlled pushing/poking) constitutes a foundational robotic skill, its learning remains challenging due to the high sensitivity to comple…
Co-Fix3D: Enhancing 3D Object Detection with Collaborative Refinement
Wenxuan Li, Qin Zou, Chi Chen +4
3D object detection in driving scenarios faces the challenge of complex road environments, which can lead to the loss or incompleteness of key features, thereby affecting perceptio…