1 paper
Jincheng Tang, Yilong Zhu, Zhengyuan Xie +2
Vision-Language-Action (VLA) models have shown remarkable promise in generalized robotic manipulation. However, their spatial generalization remains fragile. We argue that simply i…