1 paper · 1 filter
Dwip Dalal, Shivansh Patel, Chahit Jain +7
Finetuning a pretrained vision-language model (VLM) on robot demonstrations via behavior cloning (BC) has become the standard recipe for vision-language-action (VLA) policies. Howe…