1 paper
Donggeon Kim, Seungwon Jan, Hyeonjun Park +1
The reliance on language in Vision-Language-Action (VLA) models introduces ambiguity, cognitive overhead, and difficulties in precise object identification and sequential task exec…