1 paper · 1 filter
Gracjan Góral, Alicja Ziarko, Piotr Miłoś +3
We investigate the ability of Vision Language Models (VLMs) to perform visual perspective taking using a new set of visual tasks inspired by established human tests. Our approach l…