1 paper · 1 filter
Yiqian Liu, Iuliia Kotseruba, John K. Tsotsos
In this paper, we study depth perception of vision-language models (VLMs) to isolate the effects of pictorial depth cues and disentangle vision and language influences on model per…