1 paper
Yanpeng Zhao, Wentao Ding, Hongtao Li +2
A recent trend in vision-language models (VLMs) has been to enhance their spatial cognition for embodied domains. Despite progress, existing evaluations have been limited both in p…