1 paper
Alessandro Adami, Tommaso Tubaldo, Marco Todescato +2
Vision-Language Models (VLMs) have recently demonstrated strong capabilities in mapping multimodal observations to robot behaviors. However, most current approaches rely on end-to-…