1 paper · 1 filter
Yi Wang, Wendi Chen, Zimo Wen +8
Pretrained vision-language-action (VLA) policies provide strong language-conditioned manipulation knowledge, but they remain largely vision-driven and can struggle once manipulatio…