1 paper
Hokyun Im, Euijin Jeong, Andrey Kolobov +2
Vision-language-action models (VLAs) trained on large-scale robotic datasets have demonstrated strong performance on manipulation tasks, including bimanual tasks. However, because…