1 paper · 1 filter
Xueyang Zhou, Guiyao Tie, Guowen Zhang +3
Vision-Language-Action (VLA) models have advanced robotic control by enabling end-to-end decision-making directly from multimodal inputs. However, their tightly coupled architectur…