1 paper
Wenxuan Song, Ziyang Zhou, Han Zhao +7
Recent advances in Vision-Language-Action (VLA) models have enabled robotic agents to integrate multimodal understanding with action execution. However, our empirical analysis reve…