1 paper
Jess Jones, Raul Santos-Rodriguez, Sabine Hauert
Vision-language models (VLMs) have demonstrated remarkable capabilities in understanding human-object interactions, but their application to robotic systems with non-humanoid morph…