1 paper
Min Song, Yoonseong Lee, Yeonhu Seo
Vision Language Models (VLMs) have demonstrated strong capabilities in understanding visual content, yet their ability to predict where humans look on user interfaces remains unexp…