1 paper · 1 filter
Jingjing Qian, Boyao Han, Chen Shi +4
Vision-Language-Action (VLA) models achieve strong generalization in robotic manipulation but remain largely reactive and 2D-centric, making them unreliable in tasks that require p…