1 paper · 1 filter
Thomas Sievers
Vision Language Models (VLMs) enable robots to visually perceive their environment as well as the actions and characteristics of their conversation partner or humans in collaborati…