From the 1 of 4 linked papers with an AI index.
4 papers
S-squared-VLA: Decoupling Semantic and Spatial Streams in Vision-Language-Action Models for Autonomous Driving
Jianguo Yu, Rukang Wang, Duanfeng Chu +3
The paper introduces S-squared-VLA, a vision‑language‑action model that separates semantic intent reasoning from spatial geometry processing to improve low‑level control for autono…
D-MoE:Dual Disentangled Diffusion Mixture-of-Experts for Style-Controllable End-to-End Autonomous Driving
Renju Feng, Rukang Wang, Ning Xi +4
Traditional end-to-end autonomous driving frameworks frequently suffer from the "style-averaging" dilemma when trained on high-variance human demonstrations, yielding homogenized,…
LIBERO-PRO: Towards Robust and Fair Evaluation of Vision-Language-Action Models Beyond Memorization
Xueyang Zhou, Yangming Xu, Guiyao Tie +5
LIBERO has emerged as a widely adopted benchmark for evaluating Vision-Language-Action (VLA) models; however, its current training and evaluation settings are problematic, often le…
Exploring the Necessity of Reasoning in LLM-based Agent Scenarios
Xueyang Zhou, Guiyao Tie, Guowen Zhang +7
The rise of Large Reasoning Models (LRMs) signifies a paradigm shift toward advanced computational reasoning. Yet, this progress disrupts traditional agent frameworks, traditionall…