cross-domain evaluation 1peak-end rule 1temporal aggregation 1video aesthetic assessment 1vision transformer 1
From the 1 of 15 linked papers with an AI index.
Showing cs.ROShow all
2 papers · 1 filter
cs.RO2026
AutoDrive-R: Incentivizing Reasoning and Self-Reflection Capacity for VLA Model in Autonomous Driving
Zhenlong Yuan, Chengxuan Qian, Jing Tang +7
Vision-Language-Action (VLA) models in autonomous driving systems have recently demonstrated transformative potential by integrating multimodal perception with decision-making capa…
cs.RO2026
ProFocus: Proactive Perception and Focused Reasoning in Vision-and-Language Navigation
Wei Xue, Mingcheng Li, Xuecheng Wu +3
Vision-and-Language Navigation (VLN) requires agents to accurately perceive complex visual environments and reason over navigation instructions and histories. However, existing met…