4 papers
Co-Fusion4D: Spatio-temporal Collaborative Fusion for Robust 3D Object Detection
Wenxuan Li, Qin Zou, Shoubing Chen +3
In autonomous driving, 3D object detection is essential for accurate perception and reliable decision-making. However, object motion and ego-motion often induce cross-frame spatiot…
CogNav: Cognitive Process Modeling for Object Goal Navigation with LLMs
Yihan Cao, Jiazhao Zhang, Zhinan Yu +5
Object goal navigation (ObjectNav) is a fundamental task in embodied AI, requiring an agent to locate a target object in previously unseen environments. This task is particularly c…
PIN-WM: Learning Physics-INformed World Models for Non-Prehensile Manipulation
Wenxuan Li, Hang Zhao, Zhiyuan Yu +4
While non-prehensile manipulation (e.g., controlled pushing/poking) constitutes a foundational robotic skill, its learning remains challenging due to the high sensitivity to comple…
Co-Fix3D: Enhancing 3D Object Detection with Collaborative Refinement
Wenxuan Li, Qin Zou, Chi Chen +4
3D object detection in driving scenarios faces the challenge of complex road environments, which can lead to the loss or incompleteness of key features, thereby affecting perceptio…