1 paper
Hassan Jaber, Refinath S N, Luca Cagliero +2
Spatial grounding remains a key limitation of vision-language-action (VLA) systems for robotic manipulation. While current models can recognize objects and follow language instruct…