1 paper
J. de Curtò, Dayani Plasencia, Diego Sánchez +1
Vision-language models (VLMs) are increasingly used as zero-shot controllers, but successful trajectories do not necessarily show that decisions are grounded in visual input: simul…