1 paper
Minhyeok Lee, Chiyoung Kim, Chanhoe Gu +5
Vision-Language-Action (VLA) models translate natural-language commands into robot action sequences, but leading systems on the LIBERO-Plus robustness benchmark use three- to seven…