1 paper · 1 filter
Ralf Römer, Yi Zhang, Yuming Li +1
To teach robots complex manipulation tasks, a common approach is to fine-tune a pre-trained vision-language-action model (VLA) on task-specific data. However, since this recipe upd…