2 papers
cs.RO2025
PhysiAgent: An Embodied Agent Framework in Physical World
Zhihao Wang, Jianxiong Li, Jinliang Zheng +6
Vision-Language-Action (VLA) models have achieved notable success but often struggle with limited generalizations. To address this, integrating generalized Vision-Language Models (…
cs.RO2025
A Novel ViDAR Device With Visual Inertial Encoder Odometry and Reinforcement Learning-Based Active SLAM Method
Zhanhua Xin, Zhihao Wang, Shenghao Zhang +7
In the field of multi-sensor fusion for simultaneous localization and mapping (SLAM), monocular cameras and IMUs are widely used to build simple and effective visual-inertial syste…