1 paper · 1 filter
Zirun Zhou, Zhengyang Xiao, Haochuan Xu +3
Recent advances in vision-language-action (VLA) models have greatly improved embodied AI, enabling robots to follow natural language instructions and perform diverse tasks. However…