1 paper · 1 filter
Zhuoyang Zhang, Shang Yang, Qinghao Hu +5
Vision-Language-Action (VLA) models convert high-level language instructions into concrete, executable actions, a task that is especially challenging in open-world environments. We…