1 paper
Yuangong Chen, Wai Keung Wong, Jiaxing Li +2
Multimodal Large Language Models (MLLMs) show strong visual perception, yet remain limited in reasoning about space under changing viewpoints. We study this challenge as Perspectiv…