3 papers
cs.CV2026
Co-Fusion4D: Spatio-temporal Collaborative Fusion for Robust 3D Object Detection
Wenxuan Li, Qin Zou, Shoubing Chen +3
In autonomous driving, 3D object detection is essential for accurate perception and reliable decision-making. However, object motion and ego-motion often induce cross-frame spatiot…
cs.RO2026
CLASP: Closed-loop Asynchronous Spatial Perception for Open-vocabulary Desktop Object Grasping
Yiran Ling, Wenxuan Li, Siying Dong +5
Robot grasping of desktop object is widely used in intelligent manufacturing, logistics, and agriculture.Although vision-language models (VLMs) show strong potential for robotic ma…
cs.LG2025
PIN-WM: Learning Physics-INformed World Models for Non-Prehensile Manipulation
Wenxuan Li, Hang Zhao, Zhiyuan Yu +4
While non-prehensile manipulation (e.g., controlled pushing/poking) constitutes a foundational robotic skill, its learning remains challenging due to the high sensitivity to comple…