3 papers
cs.CV2026
VLA-IAP: Training-Free Visual Token Pruning via Interaction Alignment for Vision-Language-Action Models
Jintao Cheng, Haozhe Wang, Weibin Li +7
Vision-Language-Action (VLA) models have rapidly advanced embodied intelligence, enabling robots to execute complex, instruction-driven tasks. However, as model capacity and visual…
cs.RO2025
Leveraging Semantic Graphs for Efficient and Robust LiDAR SLAM
Neng Wang, Huimin Lu, Zhiqiang Zheng +3
Accurate and robust simultaneous localization and mapping (SLAM) is crucial for autonomous mobile systems, typically achieved by leveraging the geometric features of the environmen…
cs.RO2025
Image-Goal Navigation Using Refined Feature Guidance and Scene Graph Enhancement
Zhicheng Feng, Xieyuanli Chen, Chenghao Shi +4
In this paper, we introduce a novel image-goal navigation approach, named RFSG. Our focus lies in leveraging the fine-grained connections between goals, observations, and the envir…